Towards the Aha Moment of Vision-Language Models
AI & ML interests
None defined yet.
Recent Activity
View all activity
models 9
MMInstruction/Qwen2-VL-72B-Video-T3
73B • Updated • 6
MMInstruction/Giraffe
8B • Updated • 11 • 2
MMInstruction/LongVA-7B-Video-T3
8B • Updated • 6
MMInstruction/Qwen-VL-ArXivCap
Text Generation • Updated • 12 • 4
MMInstruction/Qwen-VL-ArXivQA
Text Generation • Updated • 15 • 4
MMInstruction/Silkie
Text Generation • Updated • 80 • 12
MMInstruction/YingVLM
Updated • 12 • 1
MMInstruction/YingVLM-zh
Updated • 6
MMInstruction/YingVLM-Video
Updated • 7
datasets 17
MMInstruction/stock_factors
Viewer • Updated • 48.2M • 1.87k • 4
MMInstruction/OSWorld-G
Viewer • Updated • 510 • 1.31k • 6
MMInstruction/VL-RewardBench
Viewer • Updated • 1.25k • 369 • 16
MMInstruction/Video-T3-QA
Viewer • Updated • 162k • 222 • 2
MMInstruction/SuperClevr_Val
Viewer • Updated • 5k • 80 • 1
MMInstruction/Clevr_CoGenT_TrainA_R1
Viewer • Updated • 37.8k • 155 • 48
MMInstruction/Clevr_CoGenT_TrainA_70K_Complex
Viewer • Updated • 70k • 2.56k • 8
MMInstruction/Clevr_CoGenT_ValB
Viewer • Updated • 5k • 94 • 2
MMInstruction/Clevr_CoGenT_ValA
Viewer • Updated • 5k • 1.04k • 1
MMInstruction/Clevr_CoAgent_TrainA_R1
Viewer • Updated • 2.5k • 13 • 1