HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published 2 days ago • 68
GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration Paper • 2605.31039 • Published May 29 • 46
SpatialBench: Is Your Spatial Foundation Model an All-Round Player? Paper • 2605.27367 • Published May 26 • 72
PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects Paper • 2605.21572 • Published May 20 • 55
FileGram: Grounding Agent Personalization in File-System Behavioral Traces Paper • 2604.04901 • Published Apr 6 • 40
MonoArt: Progressive Structural Reasoning for Monocular Articulated 3D Reconstruction Paper • 2603.19231 • Published Mar 19 • 37
OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Paper • 2511.10560 • Published Nov 13, 2025 • 2
4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere Paper • 2602.10094 • Published Feb 10 • 2