Edmund Raile
recallmenot
·
AI & ML interests
None yet
Organizations
None yet
Q8 crashes llama.cpp?
2
#3 opened 2 months ago
by
recallmenot
how to do MTP?
🚀 2
2
#1 opened 2 months ago
by
recallmenot
Can't wait anymore!
👍 2
2
#2 opened 3 months ago
by
Gavin-chen
please release a GGUF
#2 opened 3 months ago
by
recallmenot
question: which version of Opus-Reasoning-Distilled?
3
#1 opened 3 months ago
by
recallmenot
5.0bpw output token errors?
1
#1 opened about 1 year ago
by
recallmenot
Eval bug: asymmetric layer splitting of reduced models on multiple CUDA GPUs
1
#1 opened over 1 year ago
by
recallmenot
Eval bug: asymmetric layer splitting of reduced models on multiple CUDA GPUs
2
#3 opened over 1 year ago
by
recallmenot
Eval bug: asymmetric layer splitting of reduced models on multiple CUDA GPUs
2
#3 opened over 1 year ago
by
recallmenot