EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 1 day ago • 14
WorldLine: Action-Driven Visual Simulation for Robotic Manipulation Paper • 2609.38059 • Published 1 day ago • 11
Unfold The World: Factorize 4D Properties in Reinforcing Spatial Reasoning Paper • 2609.03729 • Published 27 days ago • 10
ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Paper • 2609.00188 • Published about 1 month ago • 54
ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Paper • 2609.00188 • Published about 1 month ago • 54
GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling Paper • 2608.29335 • Published Aug 29 • 73
SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks Paper • 2606.09669 • Published Jun 8 • 50
LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation Paper • 2605.18739 • Published May 18 • 117
NeRFLiX: High-Quality Neural View Synthesis by Learning a Degradation-Driven Inter-viewpoint MiXer Paper • 2303.06919 • Published Mar 13, 2023 • 1
Texture Generation on 3D Meshes with Point-UV Diffusion Paper • 2308.10490 • Published Aug 21, 2023 • 1
MAT: Mask-Aware Transformer for Large Hole Image Inpainting Paper • 2203.15270 • Published Mar 29, 2022
Towards Efficient and Scale-Robust Ultra-High-Definition Image Demoireing Paper • 2207.09935 • Published Jul 20, 2022
Image Inpainting via Iteratively Decoupled Probabilistic Modeling Paper • 2212.02963 • Published Dec 6, 2022
UltraPixel: Advancing Ultra-High-Resolution Image Synthesis to New Peaks Paper • 2407.02158 • Published Jul 2, 2024 • 2
ControlNeXt: Powerful and Efficient Control for Image and Video Generation Paper • 2408.06070 • Published Aug 12, 2024 • 55
Grounding-IQA: Multimodal Language Grounding Model for Image Quality Assessment Paper • 2411.17237 • Published Nov 26, 2024
CoSeR: Bridging Image and Language for Cognitive Super-Resolution Paper • 2311.16512 • Published Nov 27, 2023
Boosting Diffusion-Based Text Image Super-Resolution Model Towards Generalized Real-World Scenarios Paper • 2503.07232 • Published Mar 10, 2025