Running Agents 437 Reward Bench Leaderboard 📐 437 Explore and compare model scores on RewardBench benchmarks
Runtime error Featured 142 smolagents LLM leaderboard 🏆 142 A leaderboard for LLMs powering smolagents
microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 204k • 1.62k
Running 4.06k The Ultra-Scale Playbook 🌌 4.06k The ultimate guide to training LLM on large GPU Clusters