Request versions in other sizes.

#57
by jian2023 - opened

The 27B A3B model, or the 9B and 14B Dense models, are currently stuck at a size limit that makes them unusable. It's incredibly frustrating because you either have to accept quantization below Q3 or the speed drops to an unbearable level. Even users with regular graphics cards want to embrace AI!

Sign up or log in to comment