Hierarchical Denoising For Multi-Step Visual Reasoning Paper • 2607.15278 • Published 12 days ago • 6
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published 30 days ago • 170
Why Far Looks Up: Probing Spatial Representation in Vision-Language Models Paper • 2605.30161 • Published May 28 • 60
colabfit/Halide_Perovskite_Ion_Migration_MLFF_negative_iodide_interstitial Viewer • Updated Jun 3 • 2.22k • 161 • 1
stabilityai/stable-video-diffusion-img2vid-xt Image-to-Video • 2B • Updated Jul 10, 2024 • 210k • 3.36k
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207