Luke Alonso PRO
lukealonso
AI & ML interests
None yet
Recent Activity
updated a model 3 days ago
local-inference-lab/lil-catalog updated a model 8 days ago
local-inference-lab/Qwen3.8-Flash-Next-NVFP4 updated a model 10 days ago
local-inference-lab/Qwen3.8-27B-NVFP4-QADOrganizations
Upload chat_template.jinja
#1 opened 23 days ago
by
bullpoint
2x 6kpro
4
#2 opened about 1 month ago
by
mtcl
Lossless?
👀 1
1
#1 opened 4 months ago
by
lukealonso
Looping in OpenCode
👀 1
5
#4 opened 5 months ago
by
Jon-Nielsen
The original repository has updated some files. Does this repository need to be updated?
1
#7 opened 5 months ago
by
fanhed
Serving on two devices
3
#3 opened 5 months ago
by
shadowlilac
Will it work on 2X6000 Pros
6
#1 opened 5 months ago
by
mtcl
Why not GGUF?
#6 opened 5 months ago
by
Nerdsking
Quantization of the Model
1
#9 opened 5 months ago
by
shiva2022
Link to model and docker image
👍 1
1
#2 opened 5 months ago
by
Jon-Nielsen
Fix tool calling: support array-formatted tool content (vLLM/SGLang)
#8 opened 6 months ago
by
cudaoom
w1 not matching w3 weight scales
12
#1 opened 6 months ago
by
dareposte
RuntimeError: The size of tensor a (3072) must match the size of tensor b (6144) at non-singleton dimension 1
3
#5 opened 6 months ago
by
lianyouzao
From "Doesn't Work" to 641 tok/s: GLM-5.1 NVFP4 on 6× RTX PRO 6000 Blackwell
🔥 1
#4 opened 6 months ago
by
sakamakismile
Hopper GPU?
1
#2 opened 6 months ago
by
AndrewMatienko
Request: NVFP4 version of MiniMax-M2.5-REAP-139B (to fit on a single RTX 6000 Pro)
14
#7 opened 7 months ago
by
mondovero
Crash on first request on RTX Pro 6000 x8
👍 1
6
#3 opened 7 months ago
by
koushd
nvfp4
➕👍 2
1
#1 opened 7 months ago
by
ktsaou
VLLM error for kv weight scaling - workaround
7
#6 opened 8 months ago
by
ShaunEvansMD