zhangchenchen
zhnagchenchne
AI & ML interests
None yet
Recent Activity
liked a model 13 days ago
zhnagchenchne/Qwen3.8-27B-GPTQ-W4A16 new activity 15 days ago
nvidia/Qwen3.8-Flash-Next-NVFP4:4*5090 run nvidia/Qwen3.8-Flash-Next-NVFP4 new activity 15 days ago
primitive-ai/Qwen3.8-Flash-Next-NVFP4:can run in 4*5090?Organizations
4*5090 run nvidia/Qwen3.8-Flash-Next-NVFP4
#9 opened 15 days ago
by
zhnagchenchne
can run in 4*5090?
#3 opened 15 days ago
by
zhnagchenchne
Can it run on a configuration of 4*5090?
1
#4 opened 15 days ago
by
zhnagchenchne
Unknown vLLM environment variable detected: VLLM_PLE_MMAP
2
#1 opened 18 days ago
by
zhnagchenchne
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!
1
#3 opened about 2 months ago
by
zhnagchenchne
27B where
🔥🚀 14
9
#4 opened about 2 months ago
by
pengzhi27
call vllm or sglang?
4
#2 opened 2 months ago
by
zhnagchenchne
vllm ?
1
#2 opened 3 months ago
by
zhnagchenchne
vLLM serve for Qwen3-Omni currently only supports the thinker model.
👀 1
#8 opened 11 months ago
by
zhnagchenchne
Hiding Thinking Process
2
#25 opened about 1 year ago
by
Xzendor7
Where is the inference/vllm_tool_call.py inference/vllm_chat.py generate.py ?
2
#27 opened about 1 year ago
by
zhnagchenchne
What actually is the EOS token for this model?
4
#31 opened about 1 year ago
by
jukofyork
Keep outputting empty content without stopping
2
#3 opened about 1 year ago
by
zhnagchenchne
About AWQ
#3 opened about 1 year ago
by
zhnagchenchne
Run 1T-param on A100/H100(80G)x8 using FP4
🚀🔥 5
7
#9 opened about 1 year ago
by
ghostplant
onnx-community/Qwen3-Embedding-0.6B-ONNX
#25 opened over 1 year ago
by
zhnagchenchne
The result is problematic.
1
#3 opened over 1 year ago
by
zhnagchenchne
tokenizer_config.json
#2 opened over 1 year ago
by
zhnagchenchne
DeepSeek-Prover-V1 的升级版?
1
#13 opened over 1 year ago
by
zhnagchenchne
Calibration dataset
1
#4 opened over 1 year ago
by
AlphaGaO