AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 5 days ago • 142
yang29/gpt4-llm-cleaned-ft-qwen25-7b-sgdm_galore-lr-7e-5-r512-k200-seed42 Text Generation • 8B • Updated 4 days ago • 64 • 1
openai/whisper-large-v3-turbo Automatic Speech Recognition • 0.8B • Updated Oct 4, 2024 • 8.41M • • 3.2k
Walking in the Implicit: Interactive World Exploration via Neural Scene Representation Paper • 2606.30045 • Published 29 days ago • 8
Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models Paper • 2606.03988 • Published Jun 3 • 126
iVGR: Internalizing Visually Grounded Reasoning for MLLMs with Reinforcement Learning Paper • 2605.31096 • Published May 29 • 7
Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs Paper • 2605.30611 • Published May 28 • 252
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale Paper • 2605.27235 • Published May 26 • 9
Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation Paper • 2605.25220 • Published May 24 • 7
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207