Step-Audio-EditX โ€” MLX 4-bit

Self-contained affine 4-bit MLX bundle of Step-Audio-EditX for Apple Silicon.

This is not the CUDA/vLLM checkpoint stepfun-ai/Step-Audio-EditX-AWQ-4bit. Load it with mlx-speech.

Converted from appautomaton/step-audio-editx-8bit-mlx using MLX affine quantization (bits=4, group_size=64, mode=affine).

Component Precision
Step1 LM, flow-model, VQ02 affine int4
VQ06, HiFT, CampPlus, flow conditioner bf16 (unchanged)

Total size โ‰ˆ 2.53 GB (upstream int8 โ‰ˆ 4.4 GB).

Use

pip install mlx-speech
hf download kingfang008/step-audio-editx-4bit-mlx --local-dir ./step-audio-editx-4bit-mlx

mlx-speech tts \
  --model ./step-audio-editx-4bit-mlx \
  --reference-audio reference.wav \
  --reference-text "Transcript of the reference audio." \
  --text "New cloned speech." \
  -o output.wav

Requires an Apple Silicon Mac (M1 or later).

Downloads last month
54
Safetensors
Model size
0.6B params
Tensor type
U32
ยท
BF16
ยท
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for kingfang008/step-audio-editx-4bit-mlx

Finetuned
(3)
this model