From Pareto to Preference: Personalized Test-Time Scaling via Amortized Agentic Policy Discovery Paper • 2610.09684 • Published 5 days ago • 16
Questioning the Questions: Sustaining Self-Evolution in Reasoning Models Paper • 2610.04299 • Published 9 days ago • 70
SEER: Self-Evolving Event Reasoning and Retrieval for Time Series Forecasting Paper • 2610.04109 • Published 10 days ago • 35
Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite Paper • 2610.02826 • Published 10 days ago • 103
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 21 days ago • 227
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published Sep 10 • 657
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 28 days ago • 251
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 28 days ago • 251
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published Sep 3 • 188
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published Sep 1 • 66
Rubrics as Visual-Repair Context for Self-Evolving UI-to-Code Generation Paper • 2608.24138 • Published Aug 25 • 13
Where to Look Matters: On-Policy Self-Distillation for Long-Video Understanding Paper • 2608.25356 • Published Aug 26 • 20
OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies Paper • 2607.03723 • Published Jul 4 • 5
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Paper • 2606.03102 • Published Jun 2 • 14
Share More, Search Less: Collaborative Parallel Thinking for Efficient Test-Time Scaling Paper • 2605.27030 • Published May 26 • 29
Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use Paper • 2605.14038 • Published May 13 • 13
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Paper • 2605.08083 • Published May 8 • 71
EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents Paper • 2605.13941 • Published May 13 • 16