|
Download README.md from rththr/llamacpp-binary: direct link, hf CLI and curl.
- Browser
- Download file 705 Bytes
-
https://huggingface.co/rththr/llamacpp-binary/resolve/main/README.md
- Command line
-
hf download hf://rththr/llamacpp-binary/README.md
-
curl -L -o README.md https://huggingface.co/rththr/llamacpp-binary/resolve/main/README.md
705 Bytes
| # K2Horizon llama.cpp CUDA Binary | |
| Pre-built llama.cpp binary with CUDA support for Linux x86_64. | |
| ## Repository | |
| - Source: https://github.com/MBZUAI-IFM/llama.cpp | |
| - Branch: model/K2Horizon | |
| - Version: 0.3.0-dev | |
| - Build: 10671 | |
| - Commit: 35999d101 | |
| ## Requirements | |
| - Linux x86_64 | |
| - NVIDIA GPU with CUDA support | |
| - CUDA 12.8+ | |
| ## Files | |
| | File | Description | | |
| |------|-------------| | |
| | `llama-k2horizon-cuda-linux-x64.tar.gz` | K2Horizon llama.cpp CUDA binary | | |
| | `llama.cpp-b9628-cuda-12.8-amd64.tar.gz` | llama.cpp build b9628 for CUDA 12.8 | | |
| | `sageattention-2.2.0-cp312-cp312-linux_x86_64.whl` | SageAttention 2.2.0 pre-built wheel | | |
| ## Install | |
| ```bash | |
| tar xzf llama-k2horizon-cuda-linux-x64.tar.gz | |
| ``` | |