Li dongyang
lsm03624
AI & ML interests
None yet
Recent Activity
liked a model 3 days ago
Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot liked a model 3 days ago
Comfy-Org/MiniMax-H3 liked a model 3 days ago
MiniMaxAI/MiniMax-H3Organizations
None yet
Could you convert this Tencent model to GGUF: https://huggingface.co/tencent/Hy3 This model's capability is better than DeepSeek-V4-Flash.
👀 1
1
#8 opened about 1 month ago
by
lsm03624
DeepSeek-V4-Flash is now ready to run locally! 🐳
❤️👍 8
19
#6 opened about 1 month ago
by
danielhanchen
My output is garbled
4
#4 opened about 1 month ago
by
pliskin123
hello,The model's output is garbled.
8
#1 opened 3 months ago
by
zhao198300
FP8 seems to be broken
➕ 1
3
#1 opened about 1 month ago
by
Neiko2002
You didn't quantize it, did you? FP8 can't be the same size as BF16.
1
#1 opened about 2 months ago
by
lsm03624
怎么自我认知还是deepseek?而且好像没有做快慢思考,无法自适应控制思考长度
3
#12 opened about 2 months ago
by
user48271
模型量化的效果并不理想
5
#2 opened 6 months ago
by
mediali
The VLLM installed in this way can run this model.
#1 opened 5 months ago
by
lsm03624
Unable to run (vllm/sglang)
8
#1 opened 5 months ago
by
nfunctor
Can we perform 4-bit quantization for the awq of the Step-3.5-Flash model? The VLLM can run it.
1
#3 opened 6 months ago
by
lsm03624
the faster the inference speed becomes. Why is that?
👍 1
1
#9 opened 9 months ago
by
lsm03624
GPU: 5060Ti*4, vllm version 0.11. The following error occurred:
1
#1 opened 9 months ago
by
lsm03624
Error installing from PR branch
👍 1
14
#1 opened 9 months ago
by
DrRos
Thanks!
❤️ 2
9
#1 opened about 1 year ago
by
lightenup
Seems to be working with PR and `--jinja`
🤝❤️ 3
4
#1 opened 9 months ago
by
ubergarm