Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Paper • 2607.21848 • Published 8 days ago • 8
TableVerse: A Large-scale Tabletop Dataset with Real-world Grounded Layouts for Generalizable Manipulation Paper • 2607.21017 • Published 8 days ago • 7
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 9 days ago • 72
GRASP: GRanularity-Aware Search Policy for Agentic RAG Paper • 2607.10463 • Published 20 days ago • 9
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 9 days ago • 304
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Paper • 2607.19064 • Published 10 days ago • 73
HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis Paper • 2607.17097 • Published 12 days ago • 11
Open-AoE: An Open Egocentric Manipulation Dataset and Toolchain for Embodied Learning Paper • 2607.14183 • Published 16 days ago • 68
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune Paper • 2607.18213 • Published 11 days ago • 78
RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model Paper • 2607.17977 • Published 11 days ago • 197
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 18 days ago • 148
SPEAR: A Simulator for Photorealistic Embodied AI Research Paper • 2607.06701 • Published 24 days ago • 8
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 17 days ago • 227
Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation Paper • 2607.05382 • Published 22 days ago • 87
SynthDocBench: Controlled Benchmark for Long-Context Visual Document Understanding Paper • 2607.10400 • Published 20 days ago • 71
Towards Autonomous and Auditable Medical Imaging Model Development Paper • 2607.10522 • Published 19 days ago • 20
ABot-N1: Toward a General Visual Language Navigation Foundation Model Paper • 2607.10383 • Published 17 days ago • 102
Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation Paper • 2607.08758 • Published 22 days ago • 41
Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation Paper • 2607.07608 • Published 23 days ago • 57