QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper β’ 2609.33848 β’ Published 4 days ago β’ 35
view article Article Atom2.7m: Representation-Level Specialization for Arithmetic-Aware Small Language Models ucr-max β’ Jul 7 β’ 10