majentik's picture
Add MLX quantized model with KV cache compression
01251c6 verified