GGUF
conversational
How to use from
Docker Model Runner
docker model run hf.co/SciTools/granite:
Quick Links

Granite 4 GGUF

IBM Granite 4 models in GGUF format, for Understand's local AI. Understand serves them with its bundled server, ullama, which is built on llama.cpp.

To choose a model, see Choose a model in Understand's help. It compares every model Understand offers, with SciTools' test results and memory needs, and stays current. This page only lists the files.

File In Understand Copied unmodified from
granite-4.1-3b-UD-Q4_K_XL.gguf Granite4.1-3B, the default model unsloth/granite-4.1-3b-GGUF
granite-4.1-3b-Q4_0.gguf An earlier Granite4.1-3B download, kept for installs that use it ibm-granite/granite-4.1-3b-GGUF
granite-4.0-micro-Q4_K_M.gguf Not in the model picker unsloth/granite-4.0-micro-GGUF
granite-4.0-1b-Q4_K_M.gguf Not in the model picker ibm-granite/granite-4.0-1b-GGUF

Granite 4 is released by IBM under the Apache 2.0 license. This repository redistributes the files unmodified; it is not endorsed by IBM.

Downloads last month
506
GGUF
Model size
2B params
Architecture
granite
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support