DeepSeek-V4-Flash-0731 DeepSeek-V4-Flash-0731 GGUF quants, imatrix-calibrated. For llama.cpp / Ollama / LM Studio. 6block/DeepSeek-V4-Flash-0731-GGUF Text Generation • 284B • Updated 5 days ago • 1.74k • 1
Qwen3 All Qwen3 quantized builds: GGUF, FP8, AWQ and GPTQ. Covers 4B to 32B dense and 30B-A3B MoE. 6block/Qwen3-4B-GGUF Text Generation • 4B • Updated 10 days ago • 899 6block/Qwen3-4B-FP8 Text Generation • 4B • Updated 6 days ago • 15 6block/Qwen3-4B-AWQ Text Generation • 4B • Updated 6 days ago • 14 6block/Qwen3-4B-GPTQ Text Generation • 4B • Updated 6 days ago • 16
Kimi-K3 Moonshot Kimi-K3 GGUF quants, imatrix-calibrated. For llama.cpp / Ollama / LM Studio. 6block/Kimi-K3-GGUF Text Generation • 2.8T • Updated 6 days ago • 899 • 1
DeepSeek-V4-Flash-0731 DeepSeek-V4-Flash-0731 GGUF quants, imatrix-calibrated. For llama.cpp / Ollama / LM Studio. 6block/DeepSeek-V4-Flash-0731-GGUF Text Generation • 284B • Updated 5 days ago • 1.74k • 1
Kimi-K3 Moonshot Kimi-K3 GGUF quants, imatrix-calibrated. For llama.cpp / Ollama / LM Studio. 6block/Kimi-K3-GGUF Text Generation • 2.8T • Updated 6 days ago • 899 • 1
Qwen3 All Qwen3 quantized builds: GGUF, FP8, AWQ and GPTQ. Covers 4B to 32B dense and 30B-A3B MoE. 6block/Qwen3-4B-GGUF Text Generation • 4B • Updated 10 days ago • 899 6block/Qwen3-4B-FP8 Text Generation • 4B • Updated 6 days ago • 15 6block/Qwen3-4B-AWQ Text Generation • 4B • Updated 6 days ago • 14 6block/Qwen3-4B-GPTQ Text Generation • 4B • Updated 6 days ago • 16