Roman Ivanov
perelmanych
·
AI & ML interests
None yet
Recent Activity
new activity 17 days ago
vcruz305/DeepSeek-V4.1-Flash-GGUF:can llama.cpp run directly? new activity 23 days ago
unsloth/GLM-5.3-Flash-GGUF:All GLM-5.3-Flash quants are now available. liked a model about 1 month ago
Qwen/Qwen3.8-27B-FP8Organizations
None yet
can llama.cpp run directly?
4
#1 opened 18 days ago
by
pikaqiu888
All GLM-5.3-Flash quants are now available.
🔥 4
1
#10 opened 29 days ago
by
danielhanchen
KLD Benchmarks + why Q8_K_XL vs MXFP4 naming
❤️👍 11
6
#11 opened about 2 months ago
by
danielhanchen
Dual RTX3090 and 64GB DDR4 + 78GB Swap. Speed: ~1.78 TPS
🤝👍 4
15
#8 opened about 2 months ago
by
robert1968
Relative quant qualities
1
#2 opened 2 months ago
by
nalimc
Premature ending inside final answer after long thinking
#16 opened 2 months ago
by
perelmanych
Running with recently merged llama.cpp PR
👍 6
6
#16 opened 3 months ago
by
ubergarm
Follow up question results in complete mess
#4 opened 3 months ago
by
perelmanych
Really looking forward for 9B or 12B variants
4
#20 opened 3 months ago
by
perelmanych
Definitely interested in this one!
🚀 2
25
#1 opened 11 months ago
by
mtcl
Incorrect Model Uploaded
🤗👍 18
6
#8 opened about 1 year ago
by
noteventhrice
Qwen3 coder version
#1 opened about 1 year ago
by
perelmanych
Difference from other presets
#1 opened about 1 year ago
by
perelmanych
R1 32b is much worse than QwQ ...
22
#6 opened over 1 year ago
by
mirek190
SFT (Non-RL) distillation is this good on a sub-100B model?
3
#2 opened over 1 year ago
by
KrishnaKaasyap
IQ2_XS variant
1
#2 opened over 2 years ago
by
perelmanych
When we can expect vicuna variant of CodeLlama-2 34b model?
👍 1
#10 opened almost 3 years ago
by
perelmanych
Can't load q5_1 model
3
#1 opened about 3 years ago
by
perelmanych
Error when using with web-ui "KeyError: 'model.layers.39.self_attn.q_proj.wf1'"
❤️ 1
16
#7 opened over 3 years ago
by
TheFairyMan
Error using ooba-gooba
👍 1
39
#6 opened over 3 years ago
by
blueisbest