ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 10 days ago • 304
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 8 days ago • 149
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 17 days ago • 108
Vidu S1: A Real-Time Interactive Video Generation Model Paper • 2607.03118 • Published 28 days ago • 144
The Surprising Effectiveness of Video Diffusion Models for Hand Motion Reconstruction Paper • 2606.30308 • Published Jun 29 • 8
ABot-M0.5: Unified Mobility-and-Manipulation World Action Model Paper • 2607.00678 • Published about 1 month ago • 20
HiLo-Token: Input-Adaptive High-Low Frequency Token Compression for Efficient Image Editing Paper • 2606.13898 • Published Jun 11 • 5
Reflective Prompt Tuning through Language Model Function-Calling Paper • 2605.21781 • Published May 20 • 9
EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation Paper • 2605.23271 • Published May 22 • 82
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model Paper • 2604.20796 • Published Apr 22 • 243
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver Paper • 2604.08377 • Published Apr 9 • 295
Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability Paper • 2604.06628 • Published Apr 8 • 330
Adam's Law: Textual Frequency Law on Large Language Models Paper • 2604.02176 • Published Apr 2 • 510
Astrolabe: Steering Forward-Process Reinforcement Learning for Distilled Autoregressive Video Models Paper • 2603.17051 • Published Mar 17 • 109
SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models Paper • 2603.16859 • Published Mar 17 • 248