LLama cpp, turboquant , draft mtp 50+ t/s on low end devices AtomicChat/gemma-4-E4B-it-assistant-GGUF Text Generation • 78M • Updated Jul 23 • 1.89k • 25
1 bit llms microsoft/bitnet-b1.58-2B-4T Text Generation • 0.8B • Updated Dec 17, 2025 • 20.7k • 1.5k prism-ml/Bonsai-27B-gguf Text Generation • 27B • Updated Jul 17 • 432k • 880 prism-ml/Ternary-Bonsai-27B-gguf Text Generation • 27B • Updated 30 days ago • 634k • • 1.4k
LLama cpp, turboquant , draft mtp 50+ t/s on low end devices AtomicChat/gemma-4-E4B-it-assistant-GGUF Text Generation • 78M • Updated Jul 23 • 1.89k • 25
1 bit llms microsoft/bitnet-b1.58-2B-4T Text Generation • 0.8B • Updated Dec 17, 2025 • 20.7k • 1.5k prism-ml/Bonsai-27B-gguf Text Generation • 27B • Updated Jul 17 • 432k • 880 prism-ml/Ternary-Bonsai-27B-gguf Text Generation • 27B • Updated 30 days ago • 634k • • 1.4k