SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published 7 days ago • 155
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 7 days ago • 162
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 7 days ago • 140
SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them Paper • 2607.27703 • Published 11 days ago • 25
Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models Paper • 2508.10751 • Published Aug 14, 2025 • 29
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Paper • 2411.14062 • Published Nov 21, 2024 • 1
Does Your Reasoning Model Implicitly Know When to Stop Thinking? Paper • 2602.08354 • Published Feb 9 • 267
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger Paper • 2602.08222 • Published Feb 9 • 290
Adaptive Batch-Wise Sample Scheduling for Direct Preference Optimization Paper • 2506.17252 • Published Jun 8, 2025 • 2