ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 7 days ago • 302
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published 10 days ago • 165
sashaboguraev/pythia-1b-ppt-nca_steps1000_1b-seed1024-preserve_emb Text Generation • 1B • Updated 10 days ago • 23 • 1
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 13 days ago • 203
BadWAM: When World-Action Models Dream Right but Act Wrong Paper • 2607.15207 • Published 13 days ago • 53
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Paper • 2607.12993 • Published 15 days ago • 131
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 15 days ago • 227
MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models Paper • 2607.07673 • Published 21 days ago • 14
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF Text Generation • 1B • Updated 16 days ago • 304k • 305
RynnWorld-4D: 4D Embodied World Models for Robotic Manipulation Paper • 2607.06559 • Published 22 days ago • 95
DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams Paper • 2606.21337 • Published Jun 19 • 75
DewiBrynJones/whisper-large-v2-ft-cy-2607 Automatic Speech Recognition • 2B • Updated Jun 21 • 56 • 1