ThinkV2V: Unleashing the Reasoning Capability of MLLMs for Instruction-Guided Video Editing Paper • 2609.38541 • Published 2 days ago • 5
High Fidelity Text-Guided Music Generation and Editing via Single-Stage Flow Matching Paper • 2407.03648 • Published Jul 4, 2024 • 20
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models Paper • 2505.04921 • Published May 8, 2025 • 187