Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment
Yuhao Dong
THUdyh
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
Video-IFBench: Evaluating Instruction Following of Multimodal LLMs in Video Understanding Scenarios upvoted a paper about 1 month ago
V-Rubrics: Visual Faithfulness via Rubric-Based Reinforcement Learning authored a paper 2 months ago
PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models