Instructions to use Dhanushkumar/lora_quantize_llama_model with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Dhanushkumar/lora_quantize_llama_model with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("Dhanushkumar/lora_quantize_llama_model", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- Unsloth Desktop
Download tokenizer.json from Dhanushkumar/lora_quantize_llama_model: direct link, hf CLI and curl.
- Browser
- Download file 9.09 MB
-
https://huggingface.co/Dhanushkumar/lora_quantize_llama_model/resolve/main/tokenizer.json
- Command line
-
hf download hf://Dhanushkumar/lora_quantize_llama_model/tokenizer.json
-
curl -L -o tokenizer.json https://huggingface.co/Dhanushkumar/lora_quantize_llama_model/resolve/main/tokenizer.json
9.09 MB
File too large to display, you can check the raw version instead.