Steve Li
CHNtentes
AI & ML interests
None yet
Recent Activity
liked a model 15 days ago
prism-ml/Ternary-Bonsai-2-27B-gguf liked a model 24 days ago
deepseek-ai/DeepSeek-V4.1-Flash liked a model about 1 month ago
Qwen/Qwen3.8-Flash-NextOrganizations
None yet
ship the MTP head as a standalone mtp.safetensors (like official FP8)
π₯π 2
2
#2 opened about 2 months ago
by
NKLAR5
Pulling UD_Q8 right now, MTP questions lol
5
#3 opened 2 months ago
by
Chuck572
1-bit Kimi K3 vs Claude Opus 5 vs GPT 5.6
ππ 20
25
#12 opened 2 months ago
by
danielhanchen
Can i run this on my macbook?
ππ€― 4
5
#5 opened 2 months ago
by
hugginface536
Disappointed by inference speed compared to a usual MoE on my laptop.
π§ 1
5
#10 opened 3 months ago
by
extrabigmehdi
mmproj file so small!
7
#1 opened 3 months ago
by
CHHORVORN
You open-sourced my ass - δ½ βεΌζΊβζηε±ε§οΌ
ππ 3
3
#1 opened 4 months ago
by
JLouisBiz
Noticeable Performance Decrease
π 3
4
#23 opened 5 months ago
by
WebWeaverWraith
Can I run this model on 2x H20 141GB?
3
#1 opened 5 months ago
by
CHNtentes
Is it possible to only download the mtp gguf (<1GB one) to use with existing ggufs?
4
#3 opened 5 months ago
by
CHNtentes
Will there be small models like 12b?
ππ 5
15
#164 opened 5 months ago
by
Crownelius
Too big to run locally.
π€―π 12
21
#12 opened 5 months ago
by
Dampfinchen
ζδ»₯ζηζ―ζ··εη²ΎεΊ¦ε 樑εε€ͺε€§ε―Όθ΄ζζΆθΏζ²‘ζιεη樑εεΊζ₯
6
#96 opened 5 months ago
by
lzm1066258
May I ask if there is a deployment document?
2
#10 opened 5 months ago
by
jerryliujiawei
Parameter model.layers.15.mlp.gate_gate_up_proj.weight_scale_inv not found in params_dict
5
#3 opened 5 months ago
by
CHNtentes
ε€ͺεζΎεε¦
6
#21 opened 6 months ago
by
yukojiangjiang
These are NOT actual AWQ-quantized models.
4
#2 opened 6 months ago
by
cai-cai
larger file size for same quant
5
#4 opened 6 months ago
by
CHNtentes