K2 Horizon Collection K2 Horizon models, datasets, and supporting resources • 13 items • Updated 2 days ago • 140
Running Agents 58 Leaderboard of Smol Worldcup 📈 58 Benchmark Evaluation for Small LLMs - Leaderboard
Thinking / Reasoning Models - Reg and MOEs. Collection QwQ,DeepSeek, EXONE, DeepHermes, and others "thinking/reasoning" AIs / LLMs in regular model type, MOE (mix of experts), and Hybrid model formats. • 109 items • Updated 8 days ago • 29
The Big Benchmarks Collection Collection Gathering benchmark spaces on the hub (beyond the Open LLM Leaderboard) • 13 items • Updated Nov 18, 2024 • 271
Running 14 Reproduction: Mixed Synthetic Nearest Neighbors 🧩 14 Native mixed-anchor matrix-completion audit
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive Image-Text-to-Text • 35B • Updated Apr 17 • 865k • 3.86k