Base Model
updated
mistralai/Mistral-Small-3.1-24B-Base-2503
24B • Updated • 2.44k
• 276
Text Generation
• 0.6B • Updated • 774k
• • 199
Text Generation
• 21B • Updated • 6.75M
• • 5.11k
Text Generation
• 684B • Updated • 927k
• • 14.3k
Text Generation
• 1T • Updated • 11.5k
• 306
baidu/ERNIE-4.5-0.3B-Base-PT
Text Generation
• 0.4B • Updated • 1.69k
• 26
Text Generation
• 1B • Updated • 800k
• • 2.65k
Updated • 8.79k
• 1.17k
deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
Text Generation
• 2B • Updated • 741k
• • 1.59k
baidu/ERNIE-4.5-VL-28B-A3B-Thinking
Image-Text-to-Text
• 30B • Updated • 821
• 543
deepseek-ai/DeepSeek-R1-Zero
Text Generation
• 685B • Updated • 12.8k
• 967
Text Generation
• 9B • Updated • 16.4k
• • 111
Text Generation
• 0.4B • Updated • 28.4k
• 258
Text Generation
• 3B • Updated • 712
• 43
microsoft/Phi-4-mini-flash-reasoning
Text Generation
• 4B • Updated • 1.34k
• 285
Qwen/Qwen3-VL-2B-Instruct
Image-Text-to-Text
• 2B • Updated • 2.89M
• • 477
deepseek-ai/DeepSeek-V3.2-Exp
Text Generation
• 685B • Updated • 54.5k
• • 1.01k
tencent/Hunyuan-0.5B-Pretrain
Text Generation
• 0.5B • Updated • 329
• 10
Text Generation
• 7B • Updated • 131k
• 88