Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents Paper • 2609.17708 • Published 11 days ago • 73
Negative Self-Distillation: Learning to Reason by Avoiding Flaws Paper • 2609.11699 • Published 16 days ago • 36
ForgeWM: Progressive Causal Training for Few-Step Action-Conditioned Video World Models Paper • 2608.14022 • Published Aug 14 • 24
Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation Paper • 2608.13391 • Published Aug 13 • 20
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265
ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing Paper • 2608.04956 • Published Aug 5 • 17
Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval Paper • 2608.06060 • Published Aug 6 • 41
ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts Paper • 2607.28993 • Published Jul 31 • 8
PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents Paper • 2608.04003 • Published Aug 4 • 36