๐ In a Training Loop
Jiaqing Li
ljq34952
ยท
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO upvoted a paper about 1 month ago
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements updated a model 3 months ago
loraes/full_es_model