Commit History

GGUF: canonical leading-space chat template ({{' '}}) so default tokenization matches native exactly
0e0d72f
verified

nkthebass commited on

Fix safetensors tokenizer: fast PreTrainedTokenizerFast w/ nmt_nfkc normalizer + normalized specials + leading-space template (matches native)
096e4fe
verified

nkthebass commited on

Fix safetensors tokenizer: fast PreTrainedTokenizerFast w/ nmt_nfkc normalizer + normalized specials + leading-space template (matches native)
6d435b8
verified

nkthebass commited on

Card: document GGUF add_space_prefix=false + leading-space template fix (#23840)
9e7a2ce
verified

nkthebass commited on

Fix GGUF tokenization: add_space_prefix=false + leading-space chat template (matches native; llama.cpp #23840 workaround)
cca49d1
verified

nkthebass commited on

Card: remove inaccurate 'GGUF faithful / capacity-not-bug' note (real export tokenizer bug; robustness fix in progress)
f422736
verified

nkthebass commited on

Add prompting/clarification note: GGUF is faithful; short/empty answers on some phrasings are a 326M capacity trait, not a format bug
3dc0340
verified

nkthebass commited on

Add training process + loss curve to card
6dc4802
verified

nkthebass commited on

Add training loss curve
cdbe4be
verified

nkthebass commited on

TinyBrainBot 320M V2 Math (final5): fp16 + F16 GGUF
6e7bc72
verified

nkthebass commited on

initial commit
05baf2b
verified

nkthebass commited on