Rethinking Personalized Generation: Test-Time Alignment via Factorized Ranking Models Paper • 2609.35695 • Published 3 days ago • 10
AutoMedBench: Towards Medical AutoResearch with Agentic AI Models Paper • 2606.01961 • Published Jun 3 • 27
QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks Paper • 2605.24218 • Published May 22 • 44
VideoRLVR Collection This is the collection for VideoRLVR, including the model checkpoints and all the data used for training and evaluation. • 4 items • Updated May 20