When2Think: Learning Difficulty-Aware Length Control for Efficient Hybrid Reasoning Models Paper • 2609.19671 • Published 9 days ago • 52
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 19 days ago • 373
Negative Self-Distillation: Learning to Reason by Avoiding Flaws Paper • 2609.11699 • Published 16 days ago • 36
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets Paper • 2609.05663 • Published 22 days ago • 21
DriveZero: End-to-End Driving Beyond Human Demonstrations Paper • 2609.06055 • Published 21 days ago • 57