Instructions to use Varosa/llama-model-quantized with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Varosa/llama-model-quantized with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("Varosa/llama-model-quantized", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 135 Bytes
73ac7de | 1 2 3 4 | version https://git-lfs.github.com/spec/v1
oid sha256:8daa9615cce30c259a9555b1cc250d461d1bc69980a274b44d7eda0be78076d8
size 3791725184
|