ldt-10m / README.md

Commit History

Update card to v8 (current weights): 79,375 steps / 2.60B tokens (hits target), val 3.63. Honest eval: 40-sample sweep shows mean loop frac 0.219 (up from v7's 0.167) and 14/40 degenerate — more tokens did not buy coherence. Replaces the stale v7-only card that described the weights as 'well short of 2.6B'. (#15)
c0527a5

Compactbot commited on

Update card to v7 checkpoint (step 30000, val 3.6892 / ppl 40.01)
e9f874b
verified

Compactbot commited on

Fix over-strong 'no token loops' claim: 40-sample sweep (seeds 0-4) found 1 hard loop (seed 4, loop_frac 1.0) + several elevated-loop samples. Card now says 'occasional/rare token loops' instead of 'no token loops'. All other numbers (val 3.8943, ppl 49.12, ~308M tok) re-verified and unchanged. (#13)
9165247

Compactbot commited on

Fix v4 token-count arithmetic: v4 data build is 257.9M tok (fw 150.6M + dclm 107.3M), cumulative ~308M, ~30 tok/param, 8.4x shortfall — not the previously stated ~393M/~444M/~43.1. Verified against models/ldt-10m-v4/train.log data-build lines. (#12)
99e3ef7

Compactbot commited on

Update to v4 checkpoint: 444M tokens, val 3.8943 (ppl 49.12), still incoherent (#11)
b0a40e0

Compactbot commited on

Fix fabricated sample quote + overstatement: replace sample #1 (not in eval output) with real samples from eval_ldt10m_final.json; correct "coherent/fluency" to "grammatically structured but semantically incoherent" (#10)
451f582

Compactbot commited on

Add LDT-10M model card (#7)
1a279d1

Compactbot commited on

Add honest model card (architecture, training, undertraining status, real samples)
fc4ad9c
verified

Compactbot commited on