Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 7 days ago • 141
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published 11 days ago • 44
Ichlibitiche/csa-clinical-stage-asset-intelligence-sample Viewer • Updated 23 days ago • 614 • 111 • 1
alphaedge-ai/siglip2-so400m-patch16-512-ukr-16384 Zero-Shot Image Classification • 0.9B • Updated 24 days ago • 21 • 1
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
Securing the AI Agent: A Unified Framework for Multi-Layer Agent Red Teaming Paper • 2606.31227 • Published Jun 30 • 14
Unlocking the Visual Record of Materials Science: A Large-Scale Multimodal Dataset from Scientific Literature Paper • 2606.29667 • Published Jun 29 • 11
Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents Paper • 2605.30723 • Published May 29 • 17
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433