VABench: Measuring Embodied Spatial Intelligence through Visual Demonstrations, Active Perception, and Metric Control Paper • 2609.19554 • Published 8 days ago • 41
DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat Paper • 2609.11155 • Published 15 days ago • 28
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation Paper • 2609.08084 • Published 17 days ago • 71
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets Paper • 2609.05663 • Published 21 days ago • 21
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 24 days ago • 120
Business Arena: Benchmarking LLM Agents in a Realistic Marketplace Paper • 2608.08621 • Published Aug 9 • 22