Commit History

Model card: add IFBench and SciCode results
e52f170
verified

akkikiki commited on

Update README.md
3315e86
verified

akkikiki commited on

Model card: bundled A_log fix now follows #150
40a9f21
verified

akkikiki commited on

Fix A_log: size by num_heads, slice the zero-padded checkpoint at load
9942256
verified

akkikiki commited on

Update README.md
00d2892
verified

akkikiki commited on

Model card: state both A_log fixes without ranking them
bc01502
verified

akkikiki commited on

Model card: final GPQA Diamond result, 92.96% pass@1 avg-of-16
58a284f
verified

akkikiki commited on

Model card: A_log is zero-padded; note discussion #150 as the better fix
6ec7b0f
verified

akkikiki commited on

Update README.md
c5d5e69
verified

akkikiki commited on

Update README.md
3980e06
verified

akkikiki commited on

Update README.md
12127b7
verified

akkikiki commited on

Update README.md
def9877
verified

akkikiki commited on

Update README.md
d03e3d5
verified

akkikiki commited on

Model card: scope the precision note to the weight level, unbold none
4af146e
verified

akkikiki commited on

Model card: move the W4A16 note to the end of the Deployment section
e9efe32
verified

akkikiki commited on

Model card: drop 'Honest' from the precision note heading
cddfb2a
verified

akkikiki commited on

Rewrite model card: SGLang deployment recipe, GPQA 93.22%, W4A16 reality
9e5aec5
verified

akkikiki commited on

Fix NVFP4 quantization metadata: stale per-group mxfp4 format, glob-style exclude_modules
eb07696
verified

akkikiki commited on

Revert: Kimi-K3 uses a Python token renderer (encoding_k3.py), not a Jinja template; engines implement it natively
d426143
verified

akkikiki commited on

Add chat_template.jinja (from moonshotai/Kimi-K3 discussions/66, minja variant)
dacb247
verified

akkikiki commited on

Add quant_algo=NVFP4 (flat format required by SGLang/vLLM loaders)
78595ac
verified

akkikiki commited on

Add processor files required for SGLang/vLLM startup (multimodal arch)
3eb400c
verified

akkikiki commited on

Set base_model_relation: quantized (was showing as Finetuned)
8a65030
verified

akkikiki commited on

Add files using upload-large-folder tool
f1ac38e
verified

akkikiki commited on

Add model card: bit-exact MXFP4->NVFP4 weight-only cast, deployment status
7218095
verified

akkikiki commited on

Add files using upload-large-folder tool
bb123bd
verified

akkikiki commited on

Add files using upload-large-folder tool
3f06273
verified

akkikiki commited on

Fix config_groups to describe NVFP4 (group_size 16, E4M3 scales, tensor_group)
cae75c4
verified

akkikiki commited on

initial commit
635e490
verified

akkikiki commited on