AH22-neb's picture
Fix: dequantize attention weights (kv_b_proj, q_b_proj, etc.) - were stored as raw FP8 values
f908ee8 verified