Inference Providers
Active filters: Q8
cibernicola/FLOR-6.3B-xat-Q8_0
Text Generation
• 6B • Updated • 95
cibernicola/FLOR-1.3B-xat-Q8
Text Generation
• 1B • Updated • 39
cibernicola/FLOR-6.3B-xat-Q5_K
Text Generation
• 6B • Updated • 52
prithivMLmods/Qwen2.5-Coder-7B-Instruct-GGUF
Text Generation
• 8B • Updated • 285
• 2
prithivMLmods/Qwen2.5-Coder-7B-GGUF
Text Generation
• 8B • Updated • 194
• 3
prithivMLmods/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 123
• 4
prithivMLmods/Qwen2.5-Coder-1.5B-GGUF
Text Generation
• 2B • Updated • 624
• 5
prithivMLmods/Qwen2.5-Coder-1.5B-Instruct-GGUF
Text Generation
• 2B • Updated • 208
• 3
prithivMLmods/Qwen2.5-Coder-3B-Instruct-GGUF
Text Generation
• 3B • Updated • 248
• 5
prithivMLmods/Llama-3.2-3B-GGUF
Text Generation
• 3B • Updated • 199
• 2
harisnaeem/Phi-4-mini-instruct-GGUF-Q8
Text Generation
• 4B • Updated • 38
ykarout/llama3-deepseek_Q8
Text Generation
• 8B • Updated • 22
michelkao/Ollama-3.2-GGUF
Text Generation
• 3B • Updated • 463
SiddhJagani/gpt-oss-20b-no-think-mlx-Q8
Text Generation
• 21B • Updated • 59
• 1
0.1B • Updated • 13
Terminator278/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 102
AIconjured/embeddinggemma-300M-NVFP4-Q8-GGUF
0.3B • Updated • 263
• 2