Saad Safi
saadsafi
AI & ML interests
None yet
Recent Activity
liked a model 4 days ago
Kwaipilot/KAT-Coder-V2.5-Dev liked a model 4 days ago
bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF new activity 7 days ago
mradermacher/model_requests:vectionlabs/Salience-1.5-ProOrganizations
None yet
vectionlabs/Salience-1.5-Pro
1
#2706 opened 7 days ago
by
saadsafi
VLLM NotImplementedError in vllm/model_executor/layers/quantization/base_config.py
#1 opened 25 days ago
by
saadsafi
probably an issue with vllm 0.23.0
#2 opened about 1 month ago
by
saadsafi
MTP support
14
#3 opened 2 months ago
by
Throghar
torch RuntimeError: Shape mismatch: a.size(1) = 4096, size_k = 8192
2
#1 opened about 2 months ago
by
saadsafi
QWEN35_MTP requires nextn_predict_layers > 0.
➕👍 5
14
#2 opened 3 months ago
by
jidaigeist
Low quality code generated with latest llama.cpp
2
#1 opened 3 months ago
by
saadsafi
gemma-4-26b-a4b-it-q5_k_m
1
#1 opened 3 months ago
by
saadsafi
running "MIXED" gguf with latest llama.cpp gave this error:
1
#1 opened 3 months ago
by
saadsafi
tawkeed-sa/tawkeed-40b
2
#2069 opened 4 months ago
by
saadsafi
Intel/Qwen3.5-122B-A10B-int4-AutoRound
1
#1919 opened 5 months ago
by
saadsafi
https://huggingface.co/inceptionai/Jais-2-70B-Chat
2
#1640 opened 7 months ago
by
saadsafi
https://huggingface.co/inceptionai/Jais-2-8B-Chat
2
#1641 opened 7 months ago
by
saadsafi
bad quality code generation
1
#1 opened 7 months ago
by
saadsafi
YOYO-AI/Qwen3-30B-A3B-YOYO-V5
1
#1523 opened 9 months ago
by
saadsafi
Inference with llama.cpp + Open WebUI gives repeating `?`
4
#1 opened 9 months ago
by
whoisjeremylam
gguf size
3
#1 opened about 1 year ago
by
saadsafi
invalid value for sliding_window
2
#1 opened over 1 year ago
by
AlexPoto