Running 602 Scaling test-time compute π 602 Boost LLM answers with flexible testβtime search strategies
Running 4.05k The Ultra-Scale Playbook π 4.05k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook π 3.31k The secrets to building world-class LLMs
Running on Zero Agents Featured 455 DeepSeek OCR Demo π 455 An interactive demo for the DeepSeek-OCR model.