Architectural quantization line targeting ~13 GB with solid Q4 quality and fast RAM streaming on 16GB VRAM.
IsValorum
IsValorum
AI & ML interests
I'm tired of generic quantizations that consume too many resources when they could be even higher quality and consume much less, but nobody was doing it then I thought Why don't I do it myself? And here I am, the first person to dedicate everything to quality over quantity.
Recent Activity
new activity about 13 hours ago
IsValorum/Qwen3.6-35B-A3B-MTP-APEX-I-MiniPlus-V2.1-Abliterated-GGUF:MTP questions updated a collection about 15 hours ago
APEX-I-NanoPlus updated a model about 16 hours ago
IsValorum/Qwen3.6-35B-A3B-MTP-APEX-I-NanoPlus-GGUFOrganizations
None yet