Prism: Dynamic Sparse Attention for Native 2K Joint Video-Audio Generation Model Training Paper • 2610.05416 • Published 4 days ago • 8
Agentic Visual Generation: From Generative Models to Agentic Control Paper • 2609.06758 • Published Sep 6 • 34
Agentic Visual Generation: From Generative Models to Agentic Control Paper • 2609.06758 • Published Sep 6 • 34
VA-Judger: Reward Modeling from Human Preference Feedback for Joint Video-Audio Generation Paper • 2608.18607 • Published Aug 19 • 12
VA-Judger: Reward Modeling from Human Preference Feedback for Joint Video-Audio Generation Paper • 2608.18607 • Published Aug 19 • 12
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction Paper • 2512.16900 • Published Dec 18, 2025 • 11
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation Paper • 2508.08248 • Published Aug 11, 2025 • 27
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction Paper • 2512.16900 • Published Dec 18, 2025 • 11
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction Paper • 2512.16900 • Published Dec 18, 2025 • 11
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation Paper • 2508.08248 • Published Aug 11, 2025 • 27
MotionEditor: Editing Video Motion via Content-Aware Diffusion Paper • 2311.18830 • Published Nov 30, 2023 • 1
Implicit Temporal Modeling with Learnable Alignment for Video Recognition Paper • 2304.10465 • Published Apr 20, 2023
StableAnimator: High-Quality Identity-Preserving Human Image Animation Paper • 2411.17697 • Published Nov 26, 2024