roshniramesh 's Collections int8 llm
updated
meta-llama/Llama-Guard-3-8B-INT8
Text Generation
• 8B • Updated • 11.8k
• 38
google/gemma-7b-quant-pytorch
Text Generation
• Updated • 62
• 2
INC4AI/gpt-j-6B-int8-dynamic-inc
Text Generation
• Updated • 14
• 16
Intel/t5-small-xsum-int8-dynamic-inc
Updated • 1.74k
• 1
INC4AI/bert-base-uncased-mrpc-int8-static-inc
Text Classification
• Updated • 5
Intel/bert-large-uncased-cola-int8-inc
Text Classification
• Updated • 7
INC4AI/vit-base-patch16-224-int8-static-inc
Image Classification
• Updated • 21
• 1
INC4AI/albert-base-v2-sst2-int8-static-inc
Text Classification
• Updated • 13
Intel/roberta-base-mrpc-int8-dynamic-inc
Text Classification
• Updated • 4
INC4AI/roberta-base-mrpc-int8-static-inc
Text Classification
• Updated • 12
Intel/dynamic-minilmv2-L6-H384-squad1.1-int8-static
Question Answering
• 30.1M • Updated • 9
Intel/MiniLM-L12-H384-uncased-mrpc-int8-dynamic-inc
Text Classification
• Updated • 4
INC4AI/gpt-j-6B-int8-static-inc
Text Generation
• Updated • 16
• 9
INC4AI/gpt-j-6B-pytorch-int8-static-inc
Text Generation
• Updated • 10
Intel/bert-base-uncased-CoLA-int8-inc
Text Classification
• Updated • 11
Intel/bert-base-uncased-STS-B-int8-inc
Text Classification
• Updated • 5
INC4AI/bert-base-uncased-mrpc-int8-qat-inc
Text Classification
• Updated • 11
• 1
Intel/bert-large-uncased-rte-int8-dynamic-inc
Text Classification
• Updated • 10
Intel/bert-large-uncased-rte-int8-static-inc
Text Classification
• Updated • 12
Intel/distilbert-base-uncased-distilled-squad-int8-static-inc
Question Answering
• Updated • 1.49k
• • 5
Intel/distilbert-base-uncased-MRPC-int8-dynamic-inc
Text Classification
• Updated • 10
• 1
Intel/distilbert-base-uncased-MRPC-int8-static-inc
Text Classification
• Updated • 9
INC4AI/albert-base-v2-sst2-int8-dynamic-inc
Text Classification
• Updated • 12
Intel/albert-base-v2-MRPC-int8-inc
Text Classification
• Updated • 6
Intel/bge-small-en-v1.5-rag-int8-static
Feature Extraction
• Updated • 16
• 2
Intel/bge-base-en-v1.5-rag-int8-static
Feature Extraction
• Updated • 9
INC4AI/falcon-7b-sq-int8-inc
Text Generation
• Updated • 21
amd/Llama-3.1-8B-Instruct-w-int8-a-int8-sym-test
8B • Updated • 11.8k
RedHatAI/Llama-3.2-1B-Instruct-quantized.w8a8
Text Generation
• 1B • Updated • 35.9k
• 8
FriendliAI/Meta-Llama-3-8B-int8
Text Generation
• 8B • Updated • 5
• 1
google/gemma-7b-it-quant-pytorch
Text Generation
• Updated • 62
• 11
OpenVINO/mistral-7b-instruct-v0.1-int8-ov
Text Generation
• Updated • 25
• 1
FriendliAI/Meta-Llama-3.1-8B-Instruct-int8
Text Generation
• 8B • Updated • 16.4k
• 1
Text Generation
• 14B • Updated • 138
• 7
Text Generation
• 8B • Updated • 240
• 9
Text Generation
• 2B • Updated • 188
• 5
Qwen/Qwen1.5-1.8B-Chat-GPTQ-Int8
Text Generation
• 2B • Updated • 110
• 2
Qwen/Qwen1.5-14B-Chat-GPTQ-Int8
Text Generation
• 15B • Updated • 122
• 11
Qwen/Qwen1.5-4B-Chat-GPTQ-Int8
Text Generation
• 4B • Updated • 98
• 6
Qwen/Qwen1.5-72B-Chat-GPTQ-Int8
Text Generation
• 72B • Updated • 106
• 7
Qwen/Qwen1.5-4B-Chat-GGUF
Text Generation
• 4B • Updated • 595
• 16
Qwen/Qwen1.5-0.5B-Chat-GGUF
Text Generation
• 0.6B • Updated • 14.2k
• 35
Qwen/Qwen1.5-7B-Chat-GGUF
Text Generation
• 8B • Updated • 762
• 71
Qwen/CodeQwen1.5-7B-Chat-GGUF
Text Generation
• 7B • Updated • 1.14k
• 111
Qwen/Qwen2.5-1.5B-Instruct-GPTQ-Int8
Text Generation
• 2B • Updated • 1k
• 6
Qwen/Qwen2.5-0.5B-Instruct-GPTQ-Int8
Text Generation
• 0.5B • Updated • 712
• 10
Qwen/Qwen2.5-0.5B-Instruct-GGUF
Text Generation
• 0.6B • Updated • 150k
• 120
Qwen/Qwen2-1.5B-Instruct-GGUF
Text Generation
• 2B • Updated • 16k
• 31
Qwen/Qwen2-0.5B-Instruct-GGUF
Text Generation
• 0.5B • Updated • 10.6k
• 76
Qwen/Qwen2-7B-Instruct-GGUF
Text Generation
• 8B • Updated • 7.83k
• 180
Qwen/Qwen2-0.5B-Instruct-GPTQ-Int8
Text Generation
• 0.6B • Updated • 134
• 4
Qwen/Qwen2-1.5B-Instruct-GPTQ-Int8
Text Generation
• 2B • Updated • 129
• 4
Qwen/Qwen2-7B-Instruct-GPTQ-Int8
Text Generation
• 8B • Updated • 1.74k
• 17
Qwen/Qwen2-72B-Instruct-GPTQ-Int8
Text Generation
• 73B • Updated • 577
• 15