Commit History

Chat template: accept the 0.18 content-parts form (string form unchanged); weights, tokenizer and executor metadata byte-identical
3190c73
verified

mlboydaisuke commited on

Card: base_model_relation: quantized (so the conversion lists under the base model's Quantizations, not Finetunes)
f5c14ec
verified

mlboydaisuke commited on

Link the card back to its collection and the request box
3b658a5
verified

mlboydaisuke commited on

Add desktop LiteRT-LM CLI section (import/run/serve)
13c1bb0
verified

mlboydaisuke commited on

Card: add 'Reproduce (official tools only)' section
172e92d
verified

mlboydaisuke commited on

Add the blockwise-128 int4 recipe (reproduce with stock litert-torch)
bb4fa29
verified

mlboydaisuke commited on

Rebuild via stock litert-torch (recipe.json, no custom code); same blockwise-128 int4, GSM8K 77%
1e553b4
verified

mlboydaisuke commited on

Card: blockwise-128 + cache2048 specs (iPhone ~27 tok/s, GSM8K 77%)
24b3b76
verified

mlboydaisuke commited on

Switch to blockwise-128 int4 + cache2048 (iPhone 27 tok/s, GSM8K 77% parity)
07c5809
verified

mlboydaisuke commited on

Update card: blockwise int4 + GSM8K parity table
9f89d30
verified

mlboydaisuke commited on

Replace channelwise int4 with blockwise int4 + cache4096 (GSM8K 79% ~ bf16 75%)
5a6d84a
verified

mlboydaisuke commited on

Upload model.litertlm with huggingface_hub
247ba94
verified

mlboydaisuke commited on

Upload README.md with huggingface_hub
792b4b3
verified

mlboydaisuke commited on

initial commit
80847f8
verified

mlboydaisuke commited on