QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 5 days ago • 41
K2 Horizon Collection K2 Horizon models, datasets, and supporting resources • 24 items • Updated 1 day ago • 138
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF Image-Text-to-Text • 27B • Updated about 1 month ago • 1.68M • 1.89k
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • Aug 10 • 42
Fable 5 / Claude Code Trace Datasets Collection Public Fable 5 and Claude Code trace datasets for agent research. Check licenses and deduplicate before mixing sources. • 11 items • Updated Jun 23 • 3