On-Policy Parameter Update Direction Underlies Generalization in LLM Post-Training Paper • 2609.36659 • Published 8 days ago • 83
Video Generation Models: A Survey of Post-Training and Alignment Paper • 2610.00812 • Published 7 days ago • 60
A Missing Piece for Trustworthy AI Reviewers: From Benchmarking Rhetorical Robustness to SciCore Review Paper • 2609.39027 • Published 7 days ago • 76
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 7 days ago • 92
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 10 days ago • 567
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 11 days ago • 67
MassAlloc Attention: Let Attention Allocate Its Own Compute Paper • 2609.32712 • Published 11 days ago • 75
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published Sep 3 • 188
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published Sep 1 • 66
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published Sep 1 • 92
UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City Paper • 2608.27456 • Published Aug 27 • 88
Running Featured 853 Agent Memory Leaderboard 🧠853 Unified memory evaluation · Results expected August 12.
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published Aug 19 • 100
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements Paper • 2608.17310 • Published Aug 18 • 110
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review Paper • 2608.08975 • Published Aug 10 • 48
PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails Paper • 2607.05910 • Published Jul 7 • 31
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 58
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 60