DeepSeek-V4 Collection Sub-4-bit GGUF quants of DeepSeek-V4: Flash-0731 (seven PPL-tested tiers + DSpark draft) and Pro-0813 (1.57T, factory FP4 master). • 2 items • Updated about 21 hours ago
Qwen3 Collection Qwen3 4B to 32B in every format we ship: imatrix GGUF (seven tiers each) for llama.cpp, plus FP8, AWQ and GPTQ for vLLM. • 20 items • Updated about 21 hours ago