Running 258 The ultimate guide to RL environments: building and scaling them in the LLM era π 258 Building and scaling RL environments for LLM training
Running 119 The Eiffel Tower Llama π 119 Explore the Eiffel Tower Llama experiment with open-source models
Running on CPU Upgrade Featured 3.32k The Smol Training Playbook π 3.32k The secrets to building world-class LLMs
Running 375 LLM Embeddings Explained: A Visual and Intuitive Guide π 375 How Language Models Turn Text into Meaning, From Traditional
Running 4.07k The Ultra-Scale Playbook π 4.07k The ultimate guide to training LLM on large GPU Clusters
Running 602 Scaling test-time compute π 602 Boost LLM answers with flexible testβtime search strategies
Snowflake/snowflake-arctic-embed-m Sentence Similarity β’ 0.1B β’ Updated Dec 13, 2024 β’ 392k β’ β’ 167
Running Agents 437 Reward Bench Leaderboard π 437 Explore and compare model scores on RewardBench benchmarks
Running Featured 1.45k FineWeb: decanting the web for the finest text data at scale π· 1.45k Explore and download the FineWeb webβscale text dataset