Card: macOS/iOS 27 GA wording (beta requirement dropped; beta findings dated) 690ac16 verified mlboydaisuke commited on 25 days ago
Card metadata: library_name coreai, ensure coreai tag 8a3fb94 verified mlboydaisuke commited on Sep 6
Card: base_model_relation: quantized (so the conversion lists under the base model's Quantizations, not Finetunes) 2f66df2 verified mlboydaisuke commited on Sep 4
qwen3_8_27b_verify_s9_int4lin_d4_hpost: int4 S=9 verify + hidden out for the MTP drafter (29.0 tok/s code, lossless) aaccef9 verified mlboydaisuke commited on Aug 17
qwen3_8_27b_decode_int4lin: int4 text decoder (22.2 tok/s M4 Max, gate 15/16) 53ed45d verified mlboydaisuke commited on Aug 17
qwen3_8_27b_mtp_s9_int8hu_block32_sym: MTP S=9 replay graph (shared-KV sibling) 4d6f7de verified mlboydaisuke commited on Aug 17
qwen3_8_27b_mtp_s1_int8hu_block32_sym: MTP head as S=1 stateful drafter (int8) 577a703 verified mlboydaisuke commited on Aug 17
int4 + lossless MTP ⚡Spec: code 29.0 tok/s, free 22.2 (M4 Max) 0ea489d verified mlboydaisuke commited on Aug 17
Qwen3.8-27B: remove pf32 VL decoder (fp16 overflow on real images; pf16 replaces it) 927bbee verified mlboydaisuke commited on Aug 15
Qwen3.8-27B: VL decoder pf16 (pf32 chunk overflows fp16 content-dependently) b2086ed verified mlboydaisuke commited on Aug 15
Link the card back to its collection and the request box 8095d1a verified mlboydaisuke commited on Aug 15
Qwen3.8-27B: qwen3_8_27b_vl_decode_int8hu_block32_sym_pf32 df6b93a verified mlboydaisuke commited on Aug 15