huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
Text Generation • 284B • Updated • 482k • 175
2-bit quantization reduces the memory required to store each model weight from 16 bits down to 2 bits. This compresses the model size by roughly 8x.