Even though the accuracy seems comparable, there is significant token inflation on a per task basis
#8 opened 4 days ago
by
jakubjaniak
tool_choice: "required" causes xgrammar FSM crash / infinite hang with GLM-5.2 on vLLM 0.24.0
#7 opened 6 days ago
by
paolovic
How is the actual programming performance?
#6 opened 24 days ago
by
Artom
Some accuracy benchmark results are not as good as GLM-5.2-FP8
#5 opened about 1 month ago
by
Tianjiu
Does it support MTP?
2
#4 opened about 1 month ago
by
yz342
Which version of VLLM should I use to run this checkpoint?
3
#2 opened about 1 month ago
by
gameofdimension
Can we use this model with nvfp4 kv cache?
6
#1 opened about 1 month ago
by
positiveone