Agents-A1 Collection Agents-A1 is a Long-horizon Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. • 12 items • Updated 14 days ago • 37
VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models Paper • 2606.16140 • Published Jun 15 • 124
MLEvolve: A Self-Evolving Framework for Automated Machine Learning Algorithm Discovery Paper • 2606.06473 • Published Jun 4 • 21
ACC: Compiling Agent Trajectories for Long-Context Training Paper • 2605.21850 • Published May 21 • 61
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct Text Generation • 16B • Updated Jul 3, 2024 • 540k • 627
Running on CPU Upgrade Featured 3.25k The Smol Training Playbook 📚 3.25k The secrets to building world-class LLMs