Model Card for Model ID

this is an experimental autoround based quantization for EVE-Instruct model, using W4A16 scheme. quantized using open platypus dataset. this model is 4x times smaller and drops almost no performance(however evaluation needed)

Downloads last month
12
Safetensors
Model size
24B params
Tensor type
I64
·
I32
·
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support