Commit History

Enable thinking by default, preserve reasoning; drop max_new_tokens cap
179ee67
verified

joerowell commited on

Mark </assistant> (token 24) as special in tokenizer.json
88796b9
verified

joerowell commited on

Mark </assistant> (token 24) as special to match internal serving
6548c10
verified

joerowell commited on

Fix chat template: preserve reasoning across turns (llama.cpp --reasoning-preserve)
fac825e
verified

joerowell commited on