Base Model
updated
mistralai/Mistral-Small-3.1-24B-Base-2503
24B • Updated • 1.29k
• 272
Text Generation
• 0.6B • Updated • 730k
• 178
Text Generation
• 22B • Updated • 8.25M
• • 4.87k
Text Generation
• 685B • Updated • 9.4M
• • 13.5k
Text Generation
• 1T • Updated • 8.05k
• 306
baidu/ERNIE-4.5-0.3B-Base-PT
Text Generation
• 0.4B • Updated • 3.12k
• 26
Text Generation
• 1B • Updated • 1.67M
• 2.51k
Updated • 10.9k
• 1.14k
deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
Text Generation
• 2B • Updated • 589k
• • 1.55k
baidu/ERNIE-4.5-VL-28B-A3B-Thinking
Image-Text-to-Text
• 30B • Updated • 1.36k
• 542
deepseek-ai/DeepSeek-R1-Zero
Text Generation
• 685B • Updated • 5.58k
• 959
Text Generation
• 9B • Updated • 31.6k
• • 109
Text Generation
• 0.4B • Updated • 24.2k
• 248
Text Generation
• 3B • Updated • 809
• 42
microsoft/Phi-4-mini-flash-reasoning
Text Generation
• 4B • Updated • 908
• 281
Qwen/Qwen3-VL-2B-Instruct
Image-Text-to-Text
• 2B • Updated • 2.12M
• • 449
deepseek-ai/DeepSeek-V3.2-Exp
Text Generation
• 685B • Updated • 261k
• • 992
tencent/Hunyuan-0.5B-Pretrain
Text Generation
• 0.5B • Updated • 2.35k
• 10
Text Generation
• 7B • Updated • 125k
• 80