GLM-5.3-Flash GGUF
Placeholder repo for llama.cpp GGUF quantizations of
zai-org/GLM-5.3-Flash
(321B total / 18B active, glm5_next, native multimodal).
No GGUF files yet. Conversion needs llama.cpp glm5_next support and a machine
with enough disk/RAM; this box is too small to host the convert.
Original license: MIT.
Model tree for aj9o9/GLM-5.3-Flash-GGUF
Base model
zai-org/GLM-5.3-Flash