compactlm-5m / README.md

Commit History

Fix over-stated quality claims: model is first-sentence-coherent then loops, not "fluent/on-topic" (#9)
9323e39

Compactbot commited on

Add version note: v2 replaces earlier 6.16M word-salad build
5dbd13e
verified

Compactbot commited on

Add model card: CompactLM-5M from-scratch 4.9M-param GQA LLaMA
a6ec395
verified

Compactbot commited on

Fix Results section to report the measured val loss/ppl from the shipped checkpoint's own eval (eval_fresh.json: 3.8719 / 48.03). The previous card cited 3.8775 / 48.30 / "step 20000", but the training run diverged to NaN at step 14300 and the log died at step 16000 — step 20000 was never reached. Also correct the eval filename reference (eval_shipped.json -> eval_fresh.json, the file actually in the repo). (#8)
23f2b03

Compactbot commited on

Fix: quote verbatim samples from the shipped eval, correct over-optimistic framing
186d75e
verified

Compactbot commited on

Add model card (honest: real samples, final val 3.8775 / ppl 48.30).
c353fb6
verified

Compactbot commited on

Fix card: replace hand-written sample sentences with real output from the shipped weights, and correct the capability claims to match what the model actually produces (#7)
ed33551

Compactbot commited on

Add model card (honest: arch, data, measured val PPL 48.03, 0/15 degenerate) (#6)
cafd02f

Compactbot commited on

Add CompactLM-5M: from-scratch LLaMA-style 6.2M-param English LM on fineweb-edu
da4f145
verified

Compactbot commited on