One to More, More to One: Category-Aware Iterative Expert Training for Software Engineering Agents Paper • 2609.23377 • Published 13 days ago • 50
🧠 SmolLM3 Collection Smol, multilingual, long-context reasoner • 14 items • Updated Oct 9, 2025 • 111
YaRN: Efficient Context Window Extension of Large Language Models Paper • 2309.00071 • Published Aug 31, 2023 • 88
GraphGPT: Generative Pre-trained Graph Eulerian Transformer Paper • 2401.00529 • Published Dec 31, 2023 • 1
GraphGPT: Generative Pre-trained Graph Eulerian Transformer Paper • 2401.00529 • Published Dec 31, 2023 • 1
Qwen2.5-Coder Collection Code-specific model series based on Qwen2.5 • 38 items • Updated Mar 2 • 379
Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks