Qwen3.8-Flash-Next

#2969
by jacek2024 - opened

Could you make quants for:

https://huggingface.co/Qwen/Qwen3.8-Flash-Next

(using current llama.cpp with MTP support)

valid GGUFs are here https://huggingface.co/ggml-org/Qwen3.8-Flash-Next-GGUF but these are only Q8 and Q4

It's 1 month old and the code is new

Sign up or log in to comment