Chat2Scenic: An Iterative RAG-Based Framework for Scenario Generation in Autonomous Driving Paper • 2607.14387 • Published 24 days ago • 7
LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing Paper • 2606.26740 • Published Jun 25 • 82
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
PsychoSafe: Eliciting Psychologically-Informed Refusals in Large Language Models Paper • 2606.09697 • Published Jun 8 • 7
CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval Paper • 2605.29271 • Published May 28 • 9
Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformers Paper • 2605.06169 • Published May 7 • 238
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information Paper • 2605.11609 • Published May 12 • 196
FrameSkip: Learning from Fewer but More Informative Frames in VLA Training Paper • 2605.13757 • Published May 13 • 21
MoCapAnything V2: End-to-End Motion Capture for Arbitrary Skeletons Paper • 2604.28130 • Published Apr 30 • 24
Towards Understanding the Robustness of Sparse Autoencoders Paper • 2604.18756 • Published Apr 20 • 12
GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning Paper • 2604.02721 • Published Apr 3 • 639
General Multimodal Protein Design Enables DNA-Encoding of Chemistry Paper • 2604.05181 • Published Apr 6 • 31
FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization Paper • 2603.19835 • Published Mar 20 • 353
SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise Paper • 2602.12783 • Published Feb 13 • 246
HSImul3R: Physics-in-the-Loop Reconstruction of Simulation-Ready Human-Scene Interactions Paper • 2603.15612 • Published Mar 16 • 152
Bootstrapping Exploration with Group-Level Natural Language Feedback in Reinforcement Learning Paper • 2603.04597 • Published Mar 4 • 211