What about Q6 versions?

#12
by boldos - opened

Would it make sense to release also Q6 versions for higher quality outputs, please?
(I know we have Q8, but that one is already quite big for a 32GB GPU...)

BottleCapAI org

@boldos I just published a Q6 quant in the GGUF repo - https://huggingface.co/bottlecapai/ThinkingCap-Qwen3.6-27B-GGUF/blob/main/ThinkingCap-Qwen3.6-27B-Q6_K.gguf
We are currently running extended evals of all the quantized versions, but from the limited 1-seed temp 0.0 run it seems it sits right where it should in between Q4 and Q8. Let us know if you find any issues with it, otherwise enjoy!

klasocki changed discussion status to closed

Sign up or log in to comment