Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories Paper • 2607.15330 • Published 14 days ago • 70
Seeing the Image: Prioritizing Visual Correlation by Contrastive Alignment Paper • 2405.17871 • Published May 28, 2024 • 1
Not All Pixels Are Equal: Learning Pixel Hardness for Semantic Segmentation Paper • 2305.08462 • Published May 15, 2023 • 1