Dima Li
auf1r2
·
AI & ML interests
None yet
Recent Activity
new activity about 19 hours ago
peonist-ai/halogen-qwen3.8-flash-next:Can I create my own quantization of Qwen3.8 using `hgn` format? liked a model 5 days ago
bloomer010/Ling-3.0-flash-VL-GGUF new activity 5 days ago
bartowski/Ling-3.0-flash-VL-GGUF:Context Size is 131k, not 262kOrganizations
None yet
Can I create my own quantization of Qwen3.8 using `hgn` format?
1
#12 opened about 19 hours ago
by
auf1r2
Context Size is 131k, not 262k
#2 opened 5 days ago
by
auf1r2
Поддержка llama.cpp ?
7
#2 opened 9 days ago
by
auf1r2
Q5 or Q6 quants ?
7
#6 opened 17 days ago
by
auf1r2
DeepSeek V4.1 Flash Lite
➕👍 36
16
#11 opened 20 days ago
by
KeinNiemand
Can this technique be applied to Qwen 3.8 Flash Next?
6
#30 opened 27 days ago
by
pohnean
Any chance to get it as GGUF?
➕ 4
#2 opened 21 days ago
by
auf1r2
Just the best model 💘 as always Qwen Team is number 1 🥰 .
👍🤗 4
4
#51 opened 23 days ago
by
auf1r2
Please make it work on Strix Halo
1
#7 opened 29 days ago
by
datayoda
Is DSpark already included into these GGUFs?
👍 1
1
#1 opened 29 days ago
by
auf1r2
Is Q3 family quants in plans? 🙏
👀👍 9
#2 opened 30 days ago
by
auf1r2
Can I run this model with n-gram offloaded to SSD?
1
#2 opened about 1 month ago
by
kexar
any tips on settings?
3
#1 opened about 1 month ago
by
datayoda
4x slower than it should be? 🐢
🤗👀 5
13
#38 opened about 1 month ago
by
auf1r2
Hello! Did you try bigger size with mmap?
7
#2 opened about 1 month ago
by
auf1r2
MTP support in Q4 ?
👍 4
5
#21 opened about 1 month ago
by
dpachong
pp, is this a typo?
👍 1
2
#2 opened about 1 month ago
by
auf1r2
While waiting for Q6...
👍 1
1
#17 opened about 1 month ago
by
auf1r2
My receipe 🍜 for Strix Halo 🤖
#49 opened about 1 month ago
by
auf1r2