GLM-5.3-Flash GGUF

Placeholder repo for llama.cpp GGUF quantizations of zai-org/GLM-5.3-Flash (321B total / 18B active, glm5_next, native multimodal).

No GGUF files yet. Conversion needs llama.cpp glm5_next support and a machine with enough disk/RAM; this box is too small to host the convert.

Original license: MIT.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for aj9o9/GLM-5.3-Flash-GGUF

Quantized
(12)
this model