NVIDIA-labs OO Agents: Native Python Object-Oriented Agents Paper • 2607.20709 • Published 9 days ago • 33
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 9 days ago • 72
SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration Paper • 2607.15257 • Published 15 days ago • 71
Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published 10 days ago • 35
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published 12 days ago • 166
RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 11 days ago • 197
UniVR: Thinking in Visual Space for Unified Visual Reasoning Paper • 2607.12800 • Published 17 days ago • 32
Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget Paper • 2607.13125 • Published 13 days ago • 138
MuScriptor: An Open Model for Multi-Instrument Music Transcription Paper • 2607.08168 • Published 22 days ago • 21
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity Paper • 2607.00248 • Published Jun 30 • 32
MemLearner: Learning to Query Context memory for Video World Models Paper • 2606.31734 • Published Jun 30 • 28
Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots Paper • 2606.28133 • Published Jun 26 • 40
UnityShots: Memory-Driven Multi-Shot Audio-Video Generation with Boundary-Aware Gating Paper • 2606.21661 • Published Jun 19 • 28
EventVLA: Event-Driven Visual Evidence Memory for Long-Horizon Vision-Language-Action Policies Paper • 2606.20092 • Published Jun 18 • 6
Qwen-AgentWorld: Language World Models for General Agents Paper • 2606.24597 • Published Jun 23 • 153