An Empirical Study of Harness Design for Coding Agents Paper • 2609.20804 • Published 13 days ago • 91
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents Paper • 2609.17632 • Published 15 days ago • 45
CARDEA: Auditable Reasoning Grounded in Spatial Evidence for End-to-End Coronary Angiography Interpretation Paper • 2609.06931 • Published 23 days ago • 27
Multi-Grid Post-Training for Long-Form Multi-Shot Video Generation Paper • 2609.06373 • Published 24 days ago • 17
VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement Paper • 2609.03153 • Published 28 days ago • 15
orcarouter/Qwen3.8-Flash-Next-Uncensored-GGUF Image-Text-to-Text • 177B • Updated 19 days ago • 199k • 434
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
Decision-Metric Alignment in Latent World Models: Diagnostics and Action-Conditioned Objectives for MPC Planning Paper • 2608.18746 • Published Aug 19 • 19
TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement Paper • 2608.11951 • Published Aug 12 • 9