Anjey Sapkovski
anjeysapkovski
AI & ML interests
None yet
Recent Activity
liked a model about 1 month ago
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF new activity about 1 month ago
peculiar-ragdoll/Tiel-Coder-35B-A3B-GGUF-MTP:UD Q2_K_XL model loop in a non-tool environment liked a model about 1 month ago
peculiar-ragdoll/Tiel-Coder-35B-A3B-GGUFOrganizations
None yet
UD Q2_K_XL model loop in a non-tool environment
7
#12 opened about 1 month ago
by
anjeysapkovski
Why Q3_K for ssm.apha/beta?
11
#2 opened about 2 months ago
by
anjeysapkovski
v22.2 issue: tools calls are escaping in regular output
3
#83 opened about 2 months ago
by
anjeysapkovski
Please provide the model Qwen 3.8 35B-A5B 🙏
🔥🚀 114
8
#75 opened about 2 months ago
by
highpolygonal
Qwen3.8-27B UD-Q2_K_XL: unexpectedly slow ROCm prefill; ssm_alpha/ssm_beta are IQ1_M
🔥 1
3
#59 opened about 2 months ago
by
djtrondheim
Please release the 35b or 70b version
➕ 36
14
#7 opened about 2 months ago
by
NjProVk
ThinkingCap-Qwen3.6-35b-a3b ?
❤️ 5
7
#6 opened 3 months ago
by
Narutoouz
MTP version request
🔥❤️ 3
2
#20 opened 3 months ago
by
anjeysapkovski
Request for UD quants of the model
🚀 1
#2 opened 4 months ago
by
anjeysapkovski
Same speed on 5060 Ti as llamacpp MTP model
➕ 1
#5 opened 5 months ago
by
anjeysapkovski
Starts with 50% speedup, but speed very fast decreases
2
#21 opened 5 months ago
by
seleznyov
draft with llama.cpp?
👀 2
3
#2 opened 6 months ago
by
Schnabulator
Thank you!!
🤗 3
2
#4 opened 5 months ago
by
zrfior
Inference broken with Jan
🚀👀 4
2
#22 opened 8 months ago
by
redaihf
The generation falls into constant repetition without any good result
🔥➕ 3
16
#2 opened 9 months ago
by
ddd2r2
cool model !!
👍 1
3
#3 opened 8 months ago
by
gopi87
Check in here for tok/s and benchmarks for local gguf models
👍 1
6
#1 opened 8 months ago
by
ykarout
Why does the KV cache occupy so much GPU memory?
13
#21 opened 9 months ago
by
yyg201708
1.5b?
🔥 4
7
#3 opened 10 months ago
by
cchance27
please make a 2.1 autoround model (NT)
❤️ 2
3
#1 opened 9 months ago
by
Khatvathiren