FLUX.1-schnell β Q8_0 weight cache + auxiliary tensors
Quantized deployment artifacts for black-forest-labs/FLUX.1-schnell.
flux1-schnell-q8_0.ggufβ GGML Q8_0 tensors for the model's linear projections, stored in a GGUF container keyed by the original checkpoint tensor paths (no metadata KVs). Produced by in-memory quantization of the bf16 checkpoint with a candle-based loader.- Config and tokenizer files are copied unmodified from the base repository.
Format note: this is not a llama.cpp-compatible GGUF β tensors keep their original checkpoint names and only linear projections are quantized. Load it with a runtime that pairs the cache with the auxiliary safetensors.
Weights are redistributed under the same terms as the base model.
- Downloads last month
- 144
Hardware compatibility
Log In to add your hardware
8-bit
Inference Providers NEW
This model isn't deployed by any Inference Provider. π Ask for provider support
Model tree for DeepDavid/FLUX.1-schnell-gguf-q8
Base model
black-forest-labs/FLUX.1-schnell