MLX
sglang
quantization
zml
cuda
apple-silicon
macos
inference
eschamoe
qwen3
qwen3-moe
mixture-of-experts
Instructions to use r3lax/escha-runtime-qwen3moe with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use r3lax/escha-runtime-qwen3moe with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir escha-runtime-qwen3moe r3lax/escha-runtime-qwen3moe
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
Ctrl+K