Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published 30 days ago • 315
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition Paper • 2601.16211 • Published Jul 2 • 51
Running Agents 34 Physical AI Bench Leaderboard 🤖 34 Benchmark for Physical AI generation and understanding
naver-hyperclovax/HyperCLOVAX-SEED-Vision-Instruct-3B Text Generation • 4B • Updated Sep 16, 2025 • 3.7k • 221
stabilityai/stable-diffusion-xl-base-1.0 Text-to-Image • 3B • Updated Oct 30, 2023 • 3.86M • • 8.25k