This model was converted to FP8 format from mistralai/Mistral-Small-Instruct-2409 using the llmcompressor library by vLLM. Refer to the original model card for more details on the model.

Downloads last month
10
Safetensors
Model size
22B params
Tensor type
F16
·
F8_E4M3
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for tolgaakar/Mistral-Small-Instruct-2409-FP8-Dynamic

Quantized
(45)
this model