What profile did these models use?
#1
by bangbangkido - opened
Thanks for testing Qwen3.8, the weights are quantized with the default-q4 profile (https://github.com/basecompute/baseRT/blob/main/base-convert/profiles/default-q4.json)
Thanks for testing Qwen3.8, the weights are quantized with the default-q4 profile (https://github.com/basecompute/baseRT/blob/main/base-convert/profiles/default-q4.json)