Dipika
AI & ML interests
None yet
Recent Activity
liked a model 1 day ago
RedHatAI/Kimi-K3-FP8-BLOCK updated a collection 5 days ago
Speculator Models new activity 21 days ago
RedHatAI/diffusiongemma-26B-A4B-it-FP8-dynamic:Update non-thinking chat templateOrganizations
Update non-thinking chat template
#3 opened 27 days ago
by
agesf
Update non-thinking chat template
#3 opened 27 days ago
by
agesf
Update README.md
#1 opened about 2 months ago
by
lwilkinson
Update README.md
#1 opened about 2 months ago
by
lwilkinson
Update chat_template.jinja
#6 opened 2 months ago
by
bdellabe
Update chat_template.jinja
#8 opened 2 months ago
by
bdellabe
Update chat_template.jinja
1
#4 opened 3 months ago
by
bdellabe
Update chat_template.jinja
#5 opened 3 months ago
by
bdellabe
这个量化类型的模型,4090显卡上可以用vllm部署嘛
2
#9 opened 3 months ago
by
David-LR
Delete .eval_results
#2 opened 3 months ago
by
SaylorTwift
Error running with latest Cuda 13 SGLang
❤️ 1
2
#7 opened 3 months ago
by
souvla
Any hope for a dynamic version?
1
#1 opened 3 months ago
by
mirix
How did you manage to make a model trained in FP16 work on NVFP4, making it bigger?
1
#1 opened 3 months ago
by
yangus87
Great quant!!
12
#6 opened 3 months ago
by
tasticleeze
Regarding the correctness of the int4 quantization script
1
#5 opened 3 months ago
by
traphix
config.json ignore-list needs linear_attn patterns — model produces garbage otherwise
❤️ 1
3
#4 opened 3 months ago
by
cghart123
Creation details?
1
#3 opened 3 months ago
by
traphix
Sglang?
1
#2 opened 3 months ago
by
jpsequeira
vllm-openai:cu130-nightly Error
➕ 4
3
#1 opened 3 months ago
by
andynoodles
Update tokenizer_config.json
#3 opened 4 months ago
by
bdellabe