Commit History

tokenizer_config: inline chat_template (resolve include for GGUF conversion)
b0a9fd7
verified

joerowell commited on

Enable thinking by default, preserve reasoning; drop max_new_tokens cap
179ee67
verified

joerowell commited on

Mark </assistant> (token 24) as special in tokenizer.json
88796b9
verified

joerowell commited on

Mark </assistant> (token 24) as special to match internal serving
6548c10
verified

joerowell commited on

Fix chat template: preserve reasoning across turns (llama.cpp --reasoning-preserve)
fac825e
verified

joerowell commited on