Based on: https://huggingface.co/DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF

Found this model to be better than Qwen3.6-27B Unsloth UD-Q8_K_XL so I worked to add MTP to it.

MTP Grafted from Unsloth MTP header

  • 50-100% Speed Increase in Decode vs Original
Downloads last month
-
GGUF
Model size
0.5B params
Architecture
clip
Hardware compatibility
Log In to add your hardware

6-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Jimster480/Deck-Opus-NEO-CODE-HERE-2T-OT-Q6_K-MTP-GGUF