The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 9 days ago • 64
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 11 days ago • 71
Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 15 days ago • 57
OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation Paper • 2609.22069 • Published 14 days ago • 37
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 15 days ago • 191
Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation Paper • 2609.20744 • Published 15 days ago • 53
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models Paper • 2609.14973 • Published 18 days ago • 175
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 22 days ago • 706
FastVideo/FastVideo-FastH3-4-step-Preview-v1-VSA-DataFree Text-to-Video • 35B • Updated 27 days ago • 962k • 316
The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation Paper • 2609.02367 • Published 30 days ago • 38
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 29 days ago • 186
H3-World: Turning Language Understanding into World Control Paper • 2609.01560 • Published about 1 month ago • 53
Ring Forcing: Towards Precise Long-Term Memory for Autoregressive Video Diffusion Paper • 2608.26794 • Published Aug 27 • 16
LayerRecall: A State-Conditioned Memory Router for Long-Horizon Consistency in Video Generation Paper • 2608.28460 • Published Aug 28 • 29