Video Generation Models: A Survey of Post-Training and Alignment Paper • 2610.00812 • Published 4 days ago • 44
VidHalluc: Evaluating Temporal Hallucinations in Multimodal Large Language Models for Video Understanding Paper • 2412.03735 • Published Dec 4, 2024 • 1
VideoA11y: Method and Dataset for Accessible Video Description Paper • 2502.20480 • Published Feb 27, 2025 • 1
VideoSAVi: Self-Aligned Video Language Models without Human Supervision Paper • 2412.00624 • Published Dec 1, 2024 • 2