Timon
KeyboardMasher
AI & ML interests
None yet
Recent Activity
new activity 2 days ago
prism-ml/Ternary-Bonsai-8B-gguf:Any idea when is the PR for llama.cpp comming? new activity 3 days ago
prism-ml/Ternary-Bonsai-27B-gguf:Update metadata (tokenizer.ggml.model) new activity 3 days ago
prism-ml/Bonsai-8B-gguf:Poor performanceOrganizations
None yet
Any idea when is the PR for llama.cpp comming?
👀 3
5
#3 opened 5 months ago
by
Kendolph
Update metadata (tokenizer.ggml.model)
1
#50 opened about 2 months ago
by
Lethaulte
Poor performance
8
#6 opened 3 months ago
by
KirMas
Temperature is last for Llama.cpp
#4 opened 7 days ago
by
KeyboardMasher
Correct Samplers Order?
#64 opened 9 days ago
by
KeyboardMasher
mmproj files are missing
#1 opened 28 days ago
by
KeyboardMasher
version for 16 GB VRAM
1
#5 opened about 1 month ago
by
KeyboardMasher
Feedback
1
#1 opened about 1 month ago
by
KeyboardMasher
token_embd.weight is quantized too much
#2 opened 3 months ago
by
KeyboardMasher
Gemma 4 seems to work best with high temperature for coding
👍 1
8
#21 opened 6 months ago
by
Reverger
Recommended sampler?
4
#4 opened 6 months ago
by
mratsim
Older quants get in the way
2
#1 opened 7 months ago
by
KeyboardMasher
Error with built-in Web UI
2
#3 opened about 1 year ago
by
KeyboardMasher
Thanks for IQ4_NL
❤️ 1
#1 opened about 1 year ago
by
KeyboardMasher
128k Context GGUF, please?
4
#2 opened over 1 year ago
by
MikeNate
Update README.md
#1 opened over 1 year ago
by
KeyboardMasher
Other Imatrix quants (IQ3_XS) ?
👍 3
6
#1 opened over 1 year ago
by deleted
llama.cpp inference too slow?
3
#6 opened over 1 year ago
by
ygsun
Over 2 tok/sec agg backed by NVMe SSD on 96GB RAM + 24GB VRAM AM5 rig with llama.cpp
🔥👍 4
9
#13 opened over 1 year ago
by
ubergarm
Issue with --n-gpu-layers 5 Parameter: Model Only Running on CPU
12
#10 opened over 1 year ago
by
vuk123