RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 10 days ago • 197
SUFLECA: Scaling Up Feature Learning for CAD-to-image Alignment Paper • 2607.15058 • Published 14 days ago • 8
Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 20 days ago • 85
OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers Paper • 2607.02461 • Published 28 days ago • 41
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 177
SciOrch: Learning to Orchestrate Expert LLMs for Solving Frontier Multimodal Scientific Reasoning Tasks Paper • 2606.15872 • Published Jun 14 • 12
Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement Paper • 2605.26952 • Published May 26 • 16
Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality? Paper • 2605.22109 • Published May 21 • 171
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence Paper • 2605.12882 • Published May 13 • 274