This is the first fine tune to exceed 730 "arc-c" ("735": 144 pts higher than Qwen 3.8 27B) AND 880 ARC-E (The OpenAI, Claude and Gemini "zone of intelligence") in 8 bit and over 718 arc-c in 4 bit.
This version is called TURBO because it drastically reduces thinking tokens (by 1/2 to as high as 1/10), yet maintains output detail and quality.
In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
This repo contains both "regular" and "MTP" Neo-CODER MAX DI-MATRIX (duel imatrix) GGUF quants.
PS: There are 29 additional quant repos as of this writing too, as well NVFP4 and many more as well.
This is one of 10+ Qwen 3.8 27B at or above ARC-C of 717 (all 10 exceed all core benchmarks of Qwen 3.8, 3.6 and 3.5 27B and 35B-A3B versions) - you can see the complete project and some of the training here :
Tuned, and tweaked to match the legendary Qwen 3.6 27B FF711 (2300+ likes, 4 million+ downloads) this fine tune matches the stability and power at "arc-c" 709: (118 pts higher than Qwen 3.8 27B) (The OpenAI, Claude and Gemini "zone of intelligence") in 8 bit and 701 arc-c in 4 bit AND THIS is instruct mode - thinking/reasoning is higher.
This version is called TWIN-TURBO because it drastically reduces thinking tokens (by 1/2 to as LOW as 1/20), yet maintains output detail and quality. In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
This repo contains both "regular" and "MTP" Neo and NEO MAX GGUF quants.
BUT WE WENT FURTHER:
Now with 5 reasoning modes (2 new - UltraXhigh / Einstein), and 5 instruct modes (2 new - UltraXhigh / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level).
UPDATE: This model (and a few more) will soon have 12 reasoning and 12 instruct modes plus interactive help, model embedded system to select the best reasoning/instruct mode[s] for all use cases.
Qwen 3.5 9B - The Defiant, 27B power ; now with Qwen 3.8 Reasoning modes.
640 ARC-C for both 8bit and 4bit. Model exceeds 7 of 7 benchmarks for Qwen 3.5 9B, Qwen3.5 27B, Qwen3.6 35B-A3B, and meets Qwen 3.6 27B in some cases... and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads)
NEW - Qwen 3.8 Reasoning Modes: 2 MTP quants (Q6/Q8) Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level). Model name has "plusIQ" in the name.
(there is also a extra robust "tools" version too.)
PS: I have posted a 22k output example using "spoon" mode too.
UPDATE: This model (and a few more) will soon have 12 reasoning and 12 instruct modes plus interactive help, model embedded system to select the best reasoning/instruct mode[s] for all use cases.