INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats Paper • 2510.25602 • Published Oct 29, 2025 • 81
Running on CPU Upgrade Featured 3.25k The Smol Training Playbook 📚 3.25k The secrets to building world-class LLMs
Towards a Unified View of Large Language Model Post-Training Paper • 2509.04419 • Published Sep 4, 2025 • 77
Running 358 LLM Embeddings Explained: A Visual and Intuitive Guide 🚀 358 How Language Models Turn Text into Meaning, From Traditional
SongGen: A Single Stage Auto-regressive Transformer for Text-to-Song Generation Paper • 2502.13128 • Published Feb 18, 2025 • 41
Running Agents Featured 2.15k Wan2.1 💻 2.15k Wan: Open and Advanced Large-Scale Video Generative Models
Running 3.96k The Ultra-Scale Playbook 🌌 3.96k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade 200 LLM Hallucination Leaderboard 🚀 200 View and filter LLM hallucination leaderboard
view article Article You could have designed state of the art positional encoding FL33TW00D-HF • Nov 25, 2024 • 492