OnAnOrange/qwen3.5-9b-lite-osworld-cuagym-sft-1920x1080 Image-Text-to-Text • 10B • Updated 14 days ago • 19 • 1
alphaedge-ai/siglip2-base-patch16-512-tgk-16384 Zero-Shot Image Classification • 0.2B • Updated 16 days ago • 24 • 1
ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes Paper • 2607.04439 • Published 28 days ago • 63
WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory Paper • 2607.02517 • Published about 1 month ago • 33
FAPO: Fully Autonomous Prompt Optimization of Multi-Step LLM Pipelines Paper • 2606.19605 • Published Jun 17 • 11
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Paper • 2606.02373 • Published Jun 1 • 60
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration Paper • 2605.20025 • Published May 19 • 191
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information Paper • 2605.11609 • Published May 12 • 196
Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation Paper • 2605.19833 • Published May 19 • 137