INT8-ComfyUI / README.md
ariaotp's picture
Update README.md
ddd0d06 verified
|
Raw
History Blame Contribute Delete
1.87 kB
metadata
tags:
  - comfyui
  - int8
  - qwen2_5_vl
  - qwen3
  - qwen3_vl
  - Quantizations

For INT4 convrot W4A4 models, go to https://huggingface.co/ariaotp/int4-ConvRot-W4A4-ComfyUI instead.

These models are quantized INT8(both of simple and convrot) for ComfyUI 0.27.0 or later. They are fast in Ampere(RTX30x0)/Ada Lovelace(RTX40x0)/Blackwell(RTX50x0).

Original models that I used:

They used Starnodes Model Converter, made by Starnodes2024, quantized from BF16 models https://github.com/Starnodes2024/comfyui-starnodes-modelconverter Or quant_int8_convrot.py of ComfyUI's official ComfyUI model tools https://github.com/Comfy-Org/comfy-model-tools

License

Qwen2.5-VL, Qwen3, Qwen3-VL: Apache 2.0