GSQ-RCO-GGUF Collection Non-uniform GGUF quantizations via GSQ + RCO: per-tensor mixed precision in standard GGUF form • 2 items • Updated 10 days ago • 42
Gemma 4 QAT Collection Gemma 4 QAT (Quantization-Aware Training) for 3x less memory use and near original accuracy. • 16 items • Updated about 1 month ago • 124
APEX Quants (GGUF) Collection MoE models quantized with the APEX Quantization technique ( https://github.com/mudler/apex-quant ) • 48 items • Updated 14 days ago • 149
Qwen3.5 Collection Qwen3.5 is Qwen's new model family including Qwen3.5 Small: 0.8B, 2B, 4B, 9B and Qwen3.5 Medium: 35B-A3B, 27B, 122B-A10B and 397B-A17B. • 25 items • Updated about 1 month ago • 167
Unsloth Dynamic 2.0 Quants Collection New 2.0 version of our Dynamic GGUF + Quants. Dynamic 2.0 achieves superior accuracy & SOTA quantization performance. • 122 items • Updated 3 days ago • 834