Improving Test-Time Scaling with Adaptive Looped Transformers Paper • 2609.35748 • Published 5 days ago • 54
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 30 days ago • 104
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning Paper • 2608.14290 • Published Aug 14 • 35
Enhancing Chat Language Models by Scaling High-quality Instructional Conversations Paper • 2305.14233 • Published May 23, 2023 • 8
MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding Paper • 2501.18362 • Published Jan 30, 2025 • 26
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation Paper • 2608.02287 • Published Aug 3 • 32
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 159