Inference Providers
Active filters: Q8
cibernicola/FLOR-6.3B-xat-Q8_0
Text Generation
• 6B • Updated • 95
cibernicola/FLOR-1.3B-xat-Q8
Text Generation
• 1B • Updated • 39
cibernicola/FLOR-6.3B-xat-Q5_K
Text Generation
• 6B • Updated • 52
prithivMLmods/Qwen2.5-Coder-7B-Instruct-GGUF
Text Generation
• 8B • Updated • 266
• 2
prithivMLmods/Qwen2.5-Coder-7B-GGUF
Text Generation
• 8B • Updated • 205
• 3
prithivMLmods/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 123
• 4
prithivMLmods/Qwen2.5-Coder-1.5B-GGUF
Text Generation
• 2B • Updated • 619
• 5
prithivMLmods/Qwen2.5-Coder-1.5B-Instruct-GGUF
Text Generation
• 2B • Updated • 210
• 3
prithivMLmods/Qwen2.5-Coder-3B-Instruct-GGUF
Text Generation
• 3B • Updated • 248
• 5
prithivMLmods/Llama-3.2-3B-GGUF
Text Generation
• 3B • Updated • 188
• 2
harisnaeem/Phi-4-mini-instruct-GGUF-Q8
Text Generation
• 4B • Updated • 38
ykarout/llama3-deepseek_Q8
Text Generation
• 8B • Updated • 22
michelkao/Ollama-3.2-GGUF
Text Generation
• 3B • Updated • 487
SiddhJagani/gpt-oss-20b-no-think-mlx-Q8
Text Generation
• 21B • Updated • 63
• 1
0.1B • Updated • 13
Terminator278/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 102
AIconjured/embeddinggemma-300M-NVFP4-Q8-GGUF
0.3B • Updated • 263
• 2