RefCaptioner: Multi-Reference Image-Grounded Video Captioning Paper • 2607.28509 • Published 13 days ago • 30
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 13 days ago • 55
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 27 days ago • 106
OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning Paper • 2606.26790 • Published Jun 25 • 57
FRBNet: Revisiting Low-Light Vision through Frequency-Domain Radial Basis Network Paper • 2510.23444 • Published Oct 27, 2025 • 1
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing Paper • 2509.24900 • Published Sep 29, 2025 • 54
RealUnify: Do Unified Models Truly Benefit from Unification? A Comprehensive Benchmark Paper • 2509.24897 • Published Sep 29, 2025 • 46