Instructions to use Sculptor-AI/Ursa_Minor_Smashed with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use Sculptor-AI/Ursa_Minor_Smashed with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf Sculptor-AI/Ursa_Minor_Smashed:F32 # Run inference directly in the terminal: llama cli -hf Sculptor-AI/Ursa_Minor_Smashed:F32
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf Sculptor-AI/Ursa_Minor_Smashed:F32 # Run inference directly in the terminal: llama cli -hf Sculptor-AI/Ursa_Minor_Smashed:F32
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf Sculptor-AI/Ursa_Minor_Smashed:F32 # Run inference directly in the terminal: ./llama-cli -hf Sculptor-AI/Ursa_Minor_Smashed:F32
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf Sculptor-AI/Ursa_Minor_Smashed:F32 # Run inference directly in the terminal: ./build/bin/llama-cli -hf Sculptor-AI/Ursa_Minor_Smashed:F32
Use Docker
docker model run hf.co/Sculptor-AI/Ursa_Minor_Smashed:F32
- LM Studio
- Jan
- Ollama
How to use Sculptor-AI/Ursa_Minor_Smashed with Ollama:
ollama run hf.co/Sculptor-AI/Ursa_Minor_Smashed:F32
- Unsloth Desktop
- Docker Model Runner
How to use Sculptor-AI/Ursa_Minor_Smashed with Docker Model Runner:
docker model run hf.co/Sculptor-AI/Ursa_Minor_Smashed:F32
- Lemonade
How to use Sculptor-AI/Ursa_Minor_Smashed with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull Sculptor-AI/Ursa_Minor_Smashed:F32
Run and chat with the model
lemonade run user.Ursa_Minor_Smashed-F32
List all available models
lemonade list
- Atomic Chat
Download setup.sh from Sculptor-AI/Ursa_Minor_Smashed: direct link, hf CLI and curl.
- Browser
- Download file 1.3 kB
-
https://huggingface.co/Sculptor-AI/Ursa_Minor_Smashed/resolve/main/setup.sh
- Command line
-
hf download hf://Sculptor-AI/Ursa_Minor_Smashed/setup.sh
-
curl -L -o setup.sh https://huggingface.co/Sculptor-AI/Ursa_Minor_Smashed/resolve/main/setup.sh
1.3 kB
| # Setup script for NanoGPT inference environment | |
| echo "Setting up NanoGPT inference environment..." | |
| # Create virtual environment | |
| python -m venv venv | |
| source venv/bin/activate | |
| # Install base requirements | |
| pip install torch numpy tiktoken tqdm | |
| # Install optional dependencies | |
| echo "Installing optional dependencies..." | |
| pip install transformers # For HuggingFace integration | |
| pip install gguf # For GGUF conversion | |
| pip install matplotlib jupyter # For visualization | |
| # Verify model file exists | |
| if [[ -f "model_optimized.pt" ]]; then | |
| echo "✓ Model file found: model_optimized.pt" | |
| else | |
| echo "⚠ Warning: model_optimized.pt not found in current directory" | |
| echo "Make sure you have the model file in the same directory as this script" | |
| fi | |
| # Clone llama.cpp for GGUF support (optional) | |
| read -p "Setup llama.cpp for GGUF support? (y/n) " -n 1 -r | |
| echo | |
| if [[ $REPLY =~ ^[Yy]$ ]]; then | |
| git clone https://github.com/ggerganov/llama.cpp | |
| cd llama.cpp | |
| make | |
| cd .. | |
| echo "llama.cpp built successfully" | |
| fi | |
| echo "Setup complete!" | |
| echo "" | |
| echo "To activate environment: source venv/bin/activate" | |
| echo "To run inference: python inference.py --prompt 'Your prompt here'" | |
| echo "To run chat: python chat.py" | |
| echo "To run examples: cd examples && python basic_usage.py" |