Running on Zero Agents Featured 2.28k Qwen3-TTS Demo π 2.28k Generate speech from text with voice design, cloning, or presets
Running Featured 94 Distilling 100B+ Models 40x Faster with TRL π 94 TRL distillation for 100B+ teachers, 40x faster
Runtime error MCP Featured 127 Mage-Flow π¨ 127 Efficient native-resolution image generation and editing
Running 68 Don't Train the Model, Evolve the Harness πΏ 68 Evolving an agent's harness, not its model, on Harvey's LAB
Running 251 The ultimate guide to RL environments: building and scaling them in the LLM era π 251 Building and scaling RL environments for LLM training
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text β’ 28B β’ Updated Jul 7 β’ 21.6k β’ β’ 2.95k
Running on CPU Upgrade 281 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens π 281 Explore synthetic data experiments as a visual bookshelf
mistralai/Voxtral-Mini-4B-Realtime-2602 Automatic Speech Recognition β’ 4B β’ Updated Mar 11 β’ 1.93M β’ 998
Running Featured 1.45k FineWeb: decanting the web for the finest text data at scale π· 1.45k Explore and download the FineWeb webβscale text dataset
Running 4.06k The Ultra-Scale Playbook π 4.06k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.32k The Smol Training Playbook π 3.32k The secrets to building world-class LLMs