view article Article Qwen3.8-27B-pi: Effort-Ordered Reasoning for Agentic Coding bytkim • 2 days ago • 13
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 6 days ago • 41
K2 Horizon Collection K2 Horizon models, datasets, and supporting resources • 24 items • Updated about 10 hours ago • 137
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF Image-Text-to-Text • 27B • Updated about 1 month ago • 1.68M • 1.9k
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • Aug 10 • 42
Fable 5 / Claude Code Trace Datasets Collection Public Fable 5 and Claude Code trace datasets for agent research. Check licenses and deduplicate before mixing sources. • 11 items • Updated Jun 23 • 3