Think Before You Score: Thinking Reward Model for Visual Generation Paper • 2609.37372 • Published 4 days ago • 93
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 23 days ago • 277
RefCaptioner: Multi-Reference Image-Grounded Video Captioning Paper • 2607.28509 • Published Jul 30 • 30
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published Jul 30 • 56
MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Paper • 2607.14189 • Published Jul 15 • 31
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation Paper • 2607.14202 • Published Jul 15 • 44