Article 1 Emergent Semantics Beyond Token Embeddings: A GPT-like Transformer Learns with Frozen 16‑D Binary Token-ID Embeddings (n_embed=16)
Beyond the Parameter Monolith: Modular Language Modeling This collection is provided for reproducibility of the paper's main claim Bochkov/fem-multi-mesh-1p7b Text Generation • 2B • Updated about 23 hours ago • 491 Bochkov/modular-reasoning-0p5b-g6p5-demo Text Generation • 0.5B • Updated about 23 hours ago • 453
Do Language Models Need a Trainable Input Embedding Table? This collection is provided for reproducibility of the paper's main claim Bochkov/ab_ext_learned Text Generation • 2B • Updated about 23 hours ago • 321 Bochkov/ab_ext_binary16 Text Generation • 2B • Updated about 23 hours ago • 292 Bochkov/ab_ext_gf2 Text Generation • 2B • Updated about 23 hours ago • 251
Beyond the Parameter Monolith: Modular Language Modeling This collection is provided for reproducibility of the paper's main claim Bochkov/fem-multi-mesh-1p7b Text Generation • 2B • Updated about 23 hours ago • 491 Bochkov/modular-reasoning-0p5b-g6p5-demo Text Generation • 0.5B • Updated about 23 hours ago • 453
Do Language Models Need a Trainable Input Embedding Table? This collection is provided for reproducibility of the paper's main claim Bochkov/ab_ext_learned Text Generation • 2B • Updated about 23 hours ago • 321 Bochkov/ab_ext_binary16 Text Generation • 2B • Updated about 23 hours ago • 292 Bochkov/ab_ext_gf2 Text Generation • 2B • Updated about 23 hours ago • 251
Bochkov/llm-fix-min-affine-recoded-minimal-code-table-free Text Generation • 0.5B • Updated 3 days ago • 285
Bochkov/llm-fix-min-baseline-learned-input-table-model-classic Text Generation • 0.5B • Updated 3 days ago • 296
Bochkov/growing-transformers-model-frozen-16-bit-baseline-monolyth-181m Text Generation • 0.2B • Updated Jan 9 • 25
Bochkov/growing-transformers-model-unfrozen-baseline-monolyth-247m Text Generation • 0.2B • Updated Jan 9 • 21