GLM-4.7-Flash-CoreAI / gpu-pipelined

Commit History

Re-export with coreai-torch 0.4.1 (fixes OS 27 beta 2+ load failure)
e40bf6f
verified

mlboydaisuke commited on

remove GatherMM int8hu — superseded by gather_qmm sym8 (2.6x, same quality)
e3abac7
verified

mlboydaisuke commited on

GLM-4.7-Flash: sym8 gather_qmm bundle (2.6x, clean int8)
a3f8528
verified

mlboydaisuke commited on

GLM-4.7-Flash: int8hu decode-pipelined bundle (ship)
5d6abab
verified

mlboydaisuke commited on