RLHND: Video Foundation Models as Physically Grounded Hand Trackers for Robot Learning Paper • 2610.09455 • Published 2 days ago • 20