Instructions to use Varosa/llama-model-quantized with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Varosa/llama-model-quantized with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("Varosa/llama-model-quantized", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Xet hash:
- 8c158e6873d4eda990cce3aad5af2c7d2cf9834cd971faf4427fcf0c51c68ddc
- Size of remote file:
- 3.79 GB
- SHA256:
- 8daa9615cce30c259a9555b1cc250d461d1bc69980a274b44d7eda0be78076d8
·
Xet efficiently stores Large Files inside Git, intelligently splitting files into unique chunks and accelerating uploads and downloads. More info.