Appearance Pointers -- Multimodal Region Control of Diffusion Transformers Paper • 2607.19344 • Published 6 days ago • 4
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published 12 days ago • 19
Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts Paper • 2607.00666 • Published 26 days ago • 24
MemLearner: Learning to Query Context memory for Video World Models Paper • 2606.31734 • Published 27 days ago • 28
PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation Paper • 2606.30673 • Published Jun 25 • 12
DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects Paper • 2606.15133 • Published Jun 13 • 74
Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance Paper • 2606.19195 • Published Jun 17 • 141
Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents Paper • 2606.19704 • Published Jun 18 • 41
World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible Paper • 2606.13652 • Published Jun 11 • 16
OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data Paper • 2606.13432 • Published Jun 11 • 113
LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding Paper • 2605.27365 • Published May 26 • 146
TideGS: Scalable Training of Over One Billion 3D Gaussian Splatting Primitives via Out-of-Core Optimization Paper • 2605.20150 • Published May 19 • 7
RT-Splatting: Joint Reflection-Transmission Modeling with Gaussian Splatting Paper • 2605.18263 • Published May 18 • 9
UniT: Unified Geometry Learning with Group Autoregressive Transformer Paper • 2605.21131 • Published May 20 • 8