MiniMax-H3 Collection NVFP4 text encoder for MiniMax-H3 video generation: the Qwen3-VL encoder quantized to run in ComfyUI on a single card. • 1 item • Updated about 3 hours ago
Qwen3.8 Collection Qwen3.8-27B (dense, vision) imatrix GGUF: 18 tiers plus q8_0/f16 mmproj and an MTP draft. More Qwen3.8 sizes as they land. • 1 item • Updated about 3 hours ago
DeepSeek-V4 Collection Sub-4-bit GGUF quants of DeepSeek-V4: Flash-0731 (seven PPL-tested tiers + DSpark draft) and Pro-0813 (1.57T, factory FP4 master). • 2 items • Updated 1 day ago
Qwen3 Collection Qwen3 4B to 32B in every format we ship: imatrix GGUF (seven tiers each) for llama.cpp, plus FP8, AWQ and GPTQ for vLLM. • 20 items • Updated 1 day ago