JMV
jamiem90
Β·
AI & ML interests
None yet
Recent Activity
reacted to danielhanchen's post with π 8 days ago
Gemma 4 is now faster and much more accurate! π
Google made huge improvements to tool-calling and chat accuracy, reliability + speed.
To get fixes, re-download our updated GGUF, MLX, NVFP4 quants!
Unsloth quants: https://huggingface.co/collections/unsloth/gemma-4
Gemma 4 Guide: https://unsloth.ai/docs/models/gemma-4 liked a model 9 days ago
unsloth/gemma-4-31B-it-GGUF reacted to danielhanchen's post with π 16 days ago
Weβre releasing new Qwen3.6 quants that run 2.5Γ faster on your GPU. β‘
Qwen3.6-27B NVFP4 runs on 24GB VRAM.
35B-A3B can hit 17,561 tok/s (B200).
We also improved accuracy, tool calling, agent use, and looping.
Qwen3.6 NVFP4: https://huggingface.co/collections/unsloth/nvfp4
Guide: https://unsloth.ai/docs/models/qwen3.6#nvfp4Organizations
None yet