Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Michael Han
shimmyshimmer
673
30
394
Follow
MatrixMemories's profile picture
fiches's profile picture
HAMMALE's profile picture
329 followers
Β·
128 following
https://unsloth.ai
unslothai
shimmyshimmer
AI & ML interests
None yet
Recent Activity
new
activity
about 11 hours ago
unsloth/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF:
Low-bit quants fall back to IQ4_NL on this arch. No i/k quant support for experts on llama.cpp.
liked
a model
about 17 hours ago
unsloth/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF
liked
a model
about 17 hours ago
unsloth/Qwen3.8-2.4T-A95B-GGUF
View all activity
Organizations
shimmyshimmer
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
unsloth/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF
about 11 hours ago
Low-bit quants fall back to IQ4_NL on this arch. No i/k quant support for experts on llama.cpp.
π
π
2
2
#1 opened 1 day ago by
worthant
New activity in
unsloth/Qwen3.8-2.4T-A95B-GGUF
about 17 hours ago
1-bit Qwen3.8 GGUF output example!
π
β€οΈ
6
3
#4 opened about 18 hours ago by
danielhanchen
New activity in
unsloth/Qwen3.6-27B-NVFP4
about 1 month ago
NVFP4 vs FP8 throughput on an RTX 6000 Pro 96 GB (Blackwell) - real vLLM numbers
β€οΈ
π
3
2
#9 opened about 1 month ago by
janreges3
New activity in
unsloth/DeepSeek-V4-Flash-GGUF
about 1 month ago
Please wait for official announcement before using!
π§
π
9
6
#5 opened about 1 month ago by
danielhanchen
My output is garbled
4
#4 opened about 1 month ago by
pliskin123
DeepSeek-V4-Flash is now ready to run locally! π³
β€οΈ
π
8
19
#6 opened about 1 month ago by
danielhanchen
Lower quants?
β
1
10
#2 opened about 1 month ago by
Kelheor
New activity in
unsloth/Qwen-AgentWorld-35B-A3B-GGUF
about 2 months ago
mmproj file so small!
7
#1 opened about 2 months ago by
CHHORVORN
New activity in
unsloth/GLM-5.2-GGUF
about 2 months ago
Jinja2 prompt template fix to get GLM5.2 running in LM Studio and Unsloth Studio
5
#6 opened about 2 months ago by
Ackerka
New activity in
unsloth/Kimi-K2.7-Code-GGUF
about 2 months ago
Gibberish output IQ3_XXS
4
#4 opened about 2 months ago by
AImhotep
New activity in
unsloth/Qwen3.6-27B-MTP-GGUF
2 months ago
One question about this GGUF
1
#18 opened 3 months ago by
alexloops
the trade off is not good (new update)
π
1
4
#28 opened 3 months ago by
rosspanda0
Qwen3.6-27b-1MδΈδΈζιζ±γ
2
#35 opened 3 months ago by
kelei999999
IQ3XXS giving garbage output
2
#33 opened 3 months ago by
mbhagya
presence-penalty
5
#8 opened 3 months ago by
owao
random chinese characters issue with unsloth GGUF for Qwen3.6-27B for the pre and post MTP
4
#36 opened 3 months ago by
Malmuk1
Don't forget to add MLX too, please! ;-)
1
#4 opened 3 months ago by
egodoyca
These shoudl work on ik_llama.cpp too
π
π
9
#3 opened 3 months ago by
ubergarm
Stable MTP first release!
β€οΈ
15
11
#6 opened 3 months ago by
danielhanchen
is it work for gemma4?
π€
1
1
#9 opened 3 months ago by
koyukira
Load more