Good intelligence, even with Q4 quantitative quantification

#5
by Healy0 - opened

Thanks for sharing this model. In my tests, it demonstrates high intelligence, even with Q4 quantization. I used a math problem derived from a physics scenario for testing; while the Q4-quantized version of ThinkingCap makes errors and fails to find the correct answer, this quantized version succeeds.

Same here, I could feel the intelligence in my short time so far for it - it arrived at the RCA of an issue faster than the typical 27b and its finetunes (I've tried loads of em). The trajectory was bang on too. Havent tried a coding session with it yet (soon), but general stuff and diagnostics are bang on. I'm using Q6 though.

Same here, I could feel the intelligence in my short time so far for it - it arrived at the RCA of an issue faster than the typical 27b and its finetunes (I've tried loads of em). The trajectory was bang on too. Havent tried a coding session with it yet (soon), but general stuff and diagnostics are bang on. I'm using Q6 though.
Same here; Tested the results of different quantization settings.

I've done coding and general tasks with it, and A/B'd against vanilla Qwen. It is more measured in its thinking, less prone to "charging" into mistakes, and generally feels and acts more like Opus.

It's a really, really good merge.

The only drawback of this model is that its name is too long, making it impossible to read the full name of each quantization version in the model list. 🀣

DavidAU pinned discussion

Sign up or log in to comment