Re-export with coreai-torch 0.4.1 (fixes OS 27 beta 2+ load failure); gated token-for-token vs fp32 oracle 8dc3742 verified mlboydaisuke commited on 21 days ago
Note b2 (beta3-loadable) bundles under gpu-pipelined-b2/ fc2e644 verified mlboydaisuke commited on 24 days ago
Add b2 (beta3-loadable) bundle under gpu-pipelined-b2/lfm2_5_1_2b_instruct_decode_int8hu_block32_sym/ (b1 retained) 9b9b2f2 verified mlboydaisuke commited on 24 days ago
Add root config.json (Hub download-stats query file + framework pointer) 333ab6b verified mlboydaisuke commited on Jun 12
head-quant docs: per-block-32 absmax ship shape (per-channel = beta delegate bug; naming note) 303538f verified mlboydaisuke commited on Jun 11
gpu-pipelined int8lin bundle: 253 tok/s M4 Max / 38-39.6 iPhone 17 Pro, oracle gate 16/16 d67f006 verified mlboydaisuke commited on Jun 10