Predictive Credit: Measuring What Scientific Explanations Add to Experimental Forecasts Paper • 2610.00314 • Published 6 days ago • 16
Decentralized Master-Mind: Joint Action Refinement through Iterative Intent Denoising in Multi-Agent Pathfinding Paper • 2609.32019 • Published 10 days ago • 44
ROWBench: Do Video Models Render What the Program Specifies? Paper • 2610.02205 • Published 4 days ago • 60
OpenTumorBoard: A Real-World Benchmark of Multidisciplinary Tumor Board Discussion Trajectories Paper • 2609.32810 • Published 9 days ago • 13
DataMagic: Authoring Data Videos through Declarative Multi-Agent Orchestration Paper • 2609.33403 • Published 8 days ago • 13
The Evolution of Attention in Large Language Models: Mechanisms, Trade-offs, and Emerging Trends Paper • 2609.39661 • Published 5 days ago • 10
Beyond Dyadic Memory: Interaction-Aware Multimodal Memory with Adaptive Agentic Retrieval for Multi-Party Spoken Conversations Paper • 2609.32522 • Published 9 days ago • 81
Synthetic Pre-pretraining Survives Scale, but Not as a Grammatical Prior Paper • 2609.39827 • Published 5 days ago • 11
CorrGRPO: Correlation-Normalized GRPO for Multi-Reward Learning Paper • 2609.36820 • Published 6 days ago • 27
Make Sparse Rewards Count: Density-Aware Reward Aggregation for Multi-Reward RL Paper • 2610.00574 • Published 5 days ago • 50
LANTERN: Illuminating Hidden Mathematical Knowledge in Language Models Paper • 2609.32264 • Published 9 days ago • 45
DuoOPD: Learning from Joint Teacher-Student Outcomes for Multi-Task On-Policy Distillation Paper • 2609.33711 • Published 8 days ago • 12
Better Supervision Is Nearby: Neighborhood On-Policy Self-Distillation Paper • 2609.39687 • Published 5 days ago • 15
The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation Paper • 2609.36484 • Published 6 days ago • 465
Omni-Embed-Mini: Binding Modalities Without Forgetting via Dense Distillation Paper • 2610.02148 • Published 4 days ago • 19
On-Policy or Off-Policy Learning? A Systematic Study of Distillation Dynamics Paper • 2609.35259 • Published 7 days ago • 168
Breaking Babel: A Self-Evolving Multi-Agent System for Long-Form Subtitle Translation Paper • 2609.38660 • Published 6 days ago • 35
Persona Dosing: Calibrated Activation Steering for Graded Trait Control Paper • 2609.36388 • Published 7 days ago • 40
Transformers Stop Thinking Too Early, and a Tiny LoRA Fixes It Paper • 2609.36585 • Published 6 days ago • 64