Huihui-RadixArk-Qwen3.8-Flash-Next-abliterated-NVFP4

Community model combining RadixArk NVFP4 with changes extracted from Huihui’s abliterated GGUF. Distributed as safetensors, with RadixArk’s original MTP, vision weights and tokenizer retained.

Tested

  • TP1 / one DGX Spark: 64K context setting; basic text, code, tool-call and image checks passed.
  • TP2 / two DGX Sparks: 1M context setting with runtime YaRN ×4; basic checks and a 66K-token retrieval request passed. Full 1M input quality was not tested.

Changes were recovered from quantized GGUF weights, so this is not an exact Huihui BF16 reconstruction. Original activation scales were retained without recalibration. Comprehensive quality and refusal-removal benchmarks remain untested.

Source revisions and weight hashes

Sources and license

Qwen Community License 1.0. Credits: Qwen, RadixArk, Huihui, Unsloth, NVIDIA ModelOpt, and llama.cpp.

Downloads last month
463
Safetensors
Model size
120B params
Tensor type
BF16
·
I64
·
U8
·
F8_E4M3
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for edp1096/Huihui-RadixArk-Qwen3.8-Flash-Next-abliterated-NVFP4

Quantized
(7)
this model