view article Article Wan 3.0 Ecosystem Signals: What WanSong, Wan-Dancer, and Wan-Streamer Reveal About Alibaba's Next Video Model ResterChed • 8 days ago • 1
Meshy T2: Fast Native Mesh Generation with Flow Matching Paper • 2607.28675 • Published 9 days ago • 53
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Paper • 2607.28611 • Published 7 days ago • 20
Explicit Layer Modeling for Video Object Insertion and Layer Decomposition Paper • 2607.25802 • Published 9 days ago • 8
ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition Paper • 2607.25565 • Published 9 days ago • 65
Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation Paper • 2607.24731 • Published 10 days ago • 76
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Paper • 2607.19064 • Published 16 days ago • 76
HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enchancement Paper • 2607.18217 • Published 17 days ago • 61
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence Paper • 2607.07675 • Published 29 days ago • 64
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation Paper • 2607.14202 • Published 22 days ago • 42
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published Jul 3 • 84
AlayaWorld: Long-Horizon and Playable Video World Generation Paper • 2607.06291 • Published about 1 month ago • 91
PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space Paper • 2607.05373 • Published Jul 6 • 66
KVpop -- Key-Value Cache Compression with Predictive Online Pruning Paper • 2607.05061 • Published Jul 6 • 24
Multi-Resolution Flow Matching: Training-Free Diffusion Acceleration via Staged Sampling Paper • 2607.01642 • Published Jul 2 • 39
LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing Paper • 2606.26740 • Published Jun 25 • 82