Bunch of bad models
Sawyer Bowerman
soyrsoyr
AI & ML interests
None yet
Recent Activity
updated a dataset 3 days ago
soyrsoyr/erebus-v2-training-data published a dataset 3 days ago
soyrsoyr/erebus-v2-training-data updated a collection 4 days ago
Custom ModelsOrganizations
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 11 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 8 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 11
Gemma 4 12B Quantized
Quantized Models
Bunch of good models
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 72 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 8 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 6 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 30
Custom Models
Bunch of bad models
Quantized Models
Bunch of good models
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 11 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 8 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 11
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 72 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 8 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 6 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 30
Gemma 4 12B Quantized