WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 6 days ago • 154
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 6 days ago • 130
GAVEL: Graph World Models for Verified and Efficient Long-Horizon LLM Task Planning Paper • 2609.19315 • Published 11 days ago • 5
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 9 days ago • 147
OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation Paper • 2609.22069 • Published 9 days ago • 36
From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention Paper • 2609.21788 • Published 9 days ago • 13
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 10 days ago • 72
Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation Paper • 2609.20744 • Published 10 days ago • 53
Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model Paper • 2609.18323 • Published 11 days ago • 132
Sample Count Is Not Enough: Candidate-Generation Strategy Shapes the Energy and Performance of LLM Test-Time Scaling Paper • 2609.19499 • Published 11 days ago • 37
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 13 days ago • 50
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 12 days ago • 46
SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators Paper • 2609.09155 • Published 19 days ago • 17
AgenticGen: Reward-Guided Agentic Video Generation for Advertising Paper • 2609.09187 • Published 27 days ago • 14