ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes Paper • 2609.01740 • Published Sep 1 • 29
LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Paper • 2608.30935 • Published Aug 31 • 33
Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion Paper • 2608.19567 • Published Aug 20 • 34
TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction Paper • 2605.26115 • Published May 25 • 55
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation Paper • 2604.24764 • Published Apr 27 • 122
Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective Paper • 2604.14025 • Published Apr 15 • 15
Scalable Adaptation of 3D Geometric Foundation Models via Weak Supervision from Internet Video Paper • 2602.07891 • Published Feb 8 • 2
InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields Paper • 2601.03252 • Published Jan 6 • 104
InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields Paper • 2601.03252 • Published Jan 6 • 104
In Pursuit of Pixel Supervision for Visual Pre-training Paper • 2512.15715 • Published Dec 17, 2025 • 11