Any progress with vision on GLM 5.3 Flash ?

#44
by Dan646464 - opened

Is work on vision in progress?

By the way for those who do that for the first time -- anyone can try GLM 5.3 Flash on text only mode taking the current pull request on llama.cpp. This is very easy. The model is very good and fast even on CPU and RAM

#1. I Assume you already have Installed llama.cpp

#2. cd ~/llama.cpp

#3. Fetch the active PR branch for GLM5-Next architecture:
git fetch origin pull/27754/head:glm5next-patch
git checkout glm5next-patch

#4.Clean up the old build files and compile the updated codebase
rm -rf build
cmake -B build
cmake --build build --config Release

#5. Chose any GLM 5.3models available on Unsloth "GLM-5.3-Flash: How to Run Locally". For example chose from the list this one: 4bit UD-Q4_K_XL which needs 200GB or 2bit UD-Q2_K_XL which has 100GB size . The download will start automatically

./build/bin/llama-server -hf unsloth/GLM-5.3-GGUF:UD-Q4_K_XL

Is work on vision in progress?

Yes.

ZHANGYUXUAN-zR changed discussion status to closed

Sign up or log in to comment