Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 126
Running 67 Don't Train the Model, Evolve the Harness 🌿 67 Evolving an agent's harness, not its model, on Harvey's LAB
Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 520
Running Featured 1.45k FineWeb: decanting the web for the finest text data at scale 🍷 1.45k Explore and download the FineWeb web‑scale text dataset
Running 4.05k The Ultra-Scale Playbook 🌌 4.05k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook 📚 3.31k The secrets to building world-class LLMs
Some of the Papers I've Read Collection A few of the research papers that I've read. • 9 items • Updated Sep 21, 2025
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Paper • 2509.02547 • Published Sep 2, 2025 • 238
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification Paper • 2502.01839 • Published Feb 3, 2025 • 10
Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge Paper • 2407.19594 • Published Jul 28, 2024 • 21
view article Article Evalverse: Revolutionizing Large Language Model Evaluation with a Unified, User-Friendly Framework Yescia • May 7, 2024 • 3
view article Article Fine-tune Llama 3.1 Ultra-Efficiently with Unsloth mlabonne • Jul 29, 2024 • 375
Running on Zero Agents 165 Gemma 3 12b It 🔥 165 Chat with an AI that understands text, images, and videos