Qwen3.8-27B MTPLX Optimized Quality

⏳ Placeholder — weights pending

The official Qwen/Qwen3.8-27B release is scheduled for 2026-08-14 15:00 UTC. This build starts the moment the official weights are readable. Like/watch this repo to get it the moment it's live. License will follow the upstream Qwen3.8-27B license.

The highest-fidelity MTPLX build of Qwen3.8-27B for Apple Silicon — for when what the model says matters more than how fast it says it.

Optimized Quality is the build for long agent sessions and the hardest tasks: staying closest to the original model's behavior while still decoding multiple tokens per step through the model's native multi-token-prediction head — preserved on load, verified with exact rejection sampling, so the output distribution matches plain decoding at real sampling settings. No greedy shortcut.

How this build is composed gets decided the way every MTPLX release is: measured on the real weights, on real hardware, against the alternatives — then shipped. Recipe details land here with the artifact, not before.

Quickstart (once weights land)

brew install youssofal/mtplx/mtplx
mtplx pull Youssofal/Qwen3.8-27B-MTPLX-Optimized-Quality
mtplx run "hello" --model Youssofal/Qwen3.8-27B-MTPLX-Optimized-Quality
mtplx serve --model Youssofal/Qwen3.8-27B-MTPLX-Optimized-Quality --port 8000

mtplx serve exposes OpenAI- and Anthropic-compatible endpoints, so the model works in anything that speaks either API.

The numbers

Posted here after the build, from max-fan verified runs on real hardware, with the workload named — the same discipline as every MTPLX release. No number on this card will predate the weights.

The MTPLX family for Qwen3.8-27B

Build Best for
Bare Speed fastest short-context chat
Optimized Speed fast coding + agents
Optimized Quality (this repo) highest fidelity, long agent sessions

MTPLX on GitHub · mtplx.com

Downloads last month
-
Safetensors
Model size
8B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Youssofal/Qwen3.8-27B-MTPLX-Optimized-Quality

Base model

Qwen/Qwen3.8-27B
Quantized
(288)
this model