UltraTex: Unleashing 2K Multi-View Diffusion for 3D Texturing Paper • 2609.23169 • Published 8 days ago • 3
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 6 days ago • 152
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 6 days ago • 56
The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing Prediction Paper • 2609.18063 • Published 11 days ago • 18
Motion-Omni: End-to-End Joint Speech and Full-Body Motion for Spoken Dialogue Paper • 2609.04250 • Published 30 days ago • 44
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published 27 days ago • 67
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 27 days ago • 96
Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors Paper • 2608.00675 • Published Aug 1 • 11
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning Paper • 2601.21716 • Published Jan 29 • 13
Parallel Decoding Distillation for Fast Image and Video Generation Paper • 2607.26004 • Published Jul 28 • 17
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published Jul 26 • 127
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published Jul 15 • 22
Motion4Motion: Motion Transfer Across Subjects at Inference Paper • 2607.11644 • Published Jul 13 • 8
AlayaWorld: Long-Horizon and Playable Video World Generation Paper • 2607.06291 • Published Jul 7 • 88
SpatialAvatar-0: High-Quality 4D Head Avatar with Multi-Stage Reconstruction Paper • 2606.15659 • Published Jun 14 • 5
PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory Paper • 2606.16449 • Published Jun 15 • 7