phak
phakio
AI & ML interests
None yet
Recent Activity
new activity about 5 hours ago
GrEarl/Kimi-K3-GGUF-IQ1_S:Working Results + Quick Benchmark liked a model about 14 hours ago
GrEarl/Kimi-K3-GGUF-IQ1_S liked a model 23 days ago
meituan-longcat/LongCat-2.0Organizations
None yet
Working Results + Quick Benchmark
❤️ 5
3
#1 opened about 14 hours ago
by
pwdz
Pretty solid results with ik_llama and OpenCode!
🔥 1
11
#2 opened about 1 month ago
by
phakio
Running with recently merged llama.cpp PR
👍 6
6
#16 opened 29 days ago
by
ubergarm
Issues with UD-Q2_K_XL quant
5
#9 opened about 1 month ago
by
labhraighlep
Q2_K has endless thinking loop issue
2
#2 opened about 1 month ago
by
phakio
really awesome speeds! running at 256k context.
🔥 1
6
#11 opened 3 months ago
by
mtcl
GLM-5.2 GGUF Benchmarks!
❤️🔥 14
26
#3 opened about 1 month ago
by
danielhanchen
Thanks for the quick quants! Reccomended mmproj?
2
#1 opened about 1 month ago
by
phakio
Any plans for reviving this model with MTP support?
👍 1
3
#14 opened 2 months ago
by
phakio
Thanks for the quantizations, can we get MTP Qwen 3.5 397B GGUF?
2
#5 opened 2 months ago
by
tidjei43
Running good on full GPU offload (1x4090, 3x3090) (Multi-GPU Offload Crash Fix)
🔥 1
#2 opened 2 months ago
by
phakio
The model is working okay! (Temporary fix for stop token being ignored)
5
#1 opened 3 months ago
by
phakio
How to use MTP in GGUF?
20
#2 opened 3 months ago
by
Friedland
Running great on my Intel QYFS and DDR5 only! (CUDA gives error)
2
#1 opened 3 months ago
by
phakio
Working good on 96GB VRAM + DDR5 Setup
❤️ 1
5
#2 opened 3 months ago
by
phakio
Great model for single GPU use cases.
🔥 4
16
#1 opened 3 months ago
by
phakio
A fun, quick little model!
#1 opened 4 months ago
by
phakio
Performance on Intel QYFS, 512GB DDR5 and 96GB VRAM
👍 1
6
#3 opened 6 months ago
by
phakio
IQ3_KS on EPYC 9355 + 1x RTX 5090
🔥 2
5
#5 opened 5 months ago
by
sousekd
This model is a gem for agentic work flows.
👍 2
1
#1 opened 5 months ago
by
phakio