File size: 314 Bytes
53e44c7
 
 
 
 
df8dcf3
 
 
1
2
3
4
5
6
7
8
---
base_model:
- poolside/Laguna-S-2.1-GGUF
---

Quantized straight from the official F16 GGUF from @poolside, using their importance matrix.

`llama-quantize --imatrix /mnt/models/laguna-s-2.1.imatrix /mnt/models/laguna-s-2.1-F16.gguf /mnt/models/laguna-s-2.1-ROCMFPX/laguna-s-2.1-Q3_0_ROCMFPX.gguf Q3_0_ROCMFPX`