RayOrch: Programming and Executing Lineage-Controlled Multi-Grain Dataflows for Foundation-Model Data Preparation Paper • 2609.18703 • Published 16 days ago • 54
Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts Paper • 2609.06011 • Published 27 days ago • 18
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 11 days ago • 130
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 14 days ago • 79
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 15 days ago • 110
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 17 days ago • 46
HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness Paper • 2609.15195 • Published 18 days ago • 22
Another Blueprint In The Wall: How to Ask Frontier AI Like a Kid? Paper • 2609.14803 • Published 19 days ago • 12
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 18 days ago • 215
Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction Paper • 2609.13285 • Published 24 days ago • 81
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 22 days ago • 46
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics Paper • 2609.10712 • Published 23 days ago • 45
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets Paper • 2609.05663 • Published 28 days ago • 21
Unlocking Lossless Speedups in LLMs via Discrete Diffusion Paper • 2609.04010 • Published 29 days ago • 114
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published about 1 month ago • 122
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM Paper • 2609.04098 • Published 29 days ago • 85
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published about 1 month ago • 91