MarkChen1214/cohere-transcribe-03-2026-MLX-Mixed-3bit4bit
Automatic Speech Recognition • 0.3B • Updated • 69 • 4
Compressed ASR for Apple Silicon: Cohere Transcribe and NVIDIA Nemotron in MLX/CoreML, including multilingual and Arabic checkpoints.
Note Qwen3-ASR 1.7B mixed 4/6/8-bit MLX model; 1.30 GB with 5.51% macro WER on the 100-sample comparison.
Note Breeze-ASR-25 uniform 5-bit MLX model; 1.07 GB with 10.30% CER on all 4,976 Common Voice 16.1 zh-TW test samples and 17.7x RTFx on the complete stable-power timing lane.