Agentic RAG Evaluation: Budget Allocation Across Questions, Trajectories, and Reads Paper • 2610.05034 • Published 6 days ago • 19
LVMT: Video Mask Transformer for Long-term Video Segmentation Paper • 2609.34895 • Published 11 days ago • 19
SemanTok: Predictable Semantic Tokens for Efficient Autoregressive Video Generation Paper • 2610.00686 • Published 10 days ago • 12
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 23 days ago • 139
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published Sep 7 • 376
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets Paper • 2609.05663 • Published Sep 4 • 21
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 154
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published Sep 3 • 182
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published Aug 18 • 161
SemaPLC: A Project-Grounded, Verification-Gated Agent Harness for PLC Code Generation Paper • 2608.18565 • Published Aug 19 • 64
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 287
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
MiniWorld: Democratizing the Training of Video World Models from Scratch Paper • 2608.01127 • Published Aug 2 • 20
Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging Paper • 2608.03316 • Published Aug 4 • 26
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published Aug 3 • 161
RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems Paper • 2607.29241 • Published Jul 31 • 12
Evaluation-Verification Reward for Consistent Multi-Reference Image Editing Paper • 2607.29025 • Published Jul 31 • 18