How to use from the
Use from the
MLX library
# Download the model from the Hub
pip install huggingface_hub[hf_xet]

huggingface-cli download --local-dir Voxtral-Mini-4B-Realtime-6bit mlx-community/Voxtral-Mini-4B-Realtime-6bit

mlx-community/Voxtral-Mini-4B-Realtime-6bit

This model mlx-community/Voxtral-Mini-4B-Realtime-6bit was converted to MLX format from mistralai/Voxtral-Mini-4B-Realtime-2602 using voxmlx.

Use with voxmlx

pip install voxmlx
from voxmlx import transcribe

text = transcribe("audio.flac", model_path="mlx-community/Voxtral-Mini-4B-Realtime-6bit")
print(text)
Downloads last month
162
Safetensors
Model size
1.0B params
Tensor type
BF16
ยท
U32
ยท
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for mlx-community/Voxtral-Mini-4B-Realtime-6bit

Spaces using mlx-community/Voxtral-Mini-4B-Realtime-6bit 3