GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models Paper • 2508.06471 • Published Aug 8, 2025 • 217
view article Article Qwen3.8-27B-pi: Effort-Ordered Reasoning for Agentic Coding bytkim • 8 days ago • 16
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 11 days ago • 46
K2 Horizon Collection K2 Horizon models, datasets, and supporting resources • 13 items • Updated 1 day ago • 140
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • Aug 10 • 42
Fable 5 / Claude Code Trace Datasets Collection Public Fable 5 and Claude Code trace datasets for agent research. Check licenses and deduplicate before mixing sources. • 11 items • Updated Jun 23 • 3
SPADE: Self-Play in Adaptive Synthetic Executable Environments Paper • 2608.19197 • Published Aug 19 • 55
Nemotron-Post-Training-v3 Collection Collection of datasets used in the post-training phase of Nemotron Nano, Super, and Ultra v3. • 50 items • Updated Aug 11 • 204
Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts Paper • 2602.13367 • Published Feb 13 • 37
view article Article Building an AI-powered search engine from scratch as-cle-bert • Dec 12, 2024 • 12
Quantization-Aware Distillation for NVFP4 Inference Accuracy Recovery Paper • 2601.20088 • Published Jan 27 • 4
view article Article AutoThink: Adaptive Reasoning for Large Language Models codelion • May 27, 2025 • 9