SGLabs/Qwen3.6-35B-A3B-Pym-Q2-MTP
Text Generation • 36B • Updated • 873 • 1
Aggressive mixed-precision GGUF quants — shrink models to a fraction of their size, keep their edge and speed. Runs on llama.cpp (AMD/Apple too).
Note First Pym release: ~13GB Q2-MTP quant of Qwen3.6-35B-A3B; MTP head preserved.