talkie-13b Collection talkie-1930-13b is a vintage language model trained on pre-1931 English-language text. See https://github.com/talkie-lm/talkie to run talkie. • 3 items • Updated 16 days ago • 46
Gemma 4 Collection Gemma 4 is Google's new model family including including E2B, E4B, 26B-A4B, and 31B. • 28 items • Updated 15 days ago • 175
view article Article KV Caching Explained: Optimizing Transformer Inference Efficiency Jan 30, 2025 • 318