Text Generation
MLX
Safetensors
English
Chinese
llama
01-ai
4-bit precision
4bit
apple-silicon
chat
code
code-generation
coding
conversational
edge-ai
instruct
local-llm
m1
m2
m3
m4
mac
mac-mini
mac-studio
macbook-air
macbook-pro
macos
metal
mlx-community
mlx-lm
no-cloud
offline
on-device
outlier
outlier-app
private
private-ai
quantized
yi
yi-coder
Instructions to use Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit with MLX:
# Make sure mlx-lm is installed # pip install --upgrade mlx-lm # Generate text with mlx-lm from mlx_lm import load, generate model, tokenizer = load("Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit") prompt = "Write a story about Einstein" messages = [{"role": "user", "content": prompt}] prompt = tokenizer.apply_chat_template( messages, add_generation_prompt=True ) text = generate(model, tokenizer, prompt=prompt, verbose=True) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- MLX LM
How to use Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit with MLX LM:
Generate or start a chat session
# Install MLX LM uv tool install mlx-lm # Interactive chat REPL mlx_lm.chat --model "Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit"
Run an OpenAI-compatible server
# Install MLX LM uv tool install mlx-lm # Start the server mlx_lm.server --model "Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit" # Calling the OpenAI-compatible server with curl curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit", "messages": [ {"role": "user", "content": "Hello"} ] }'
| license: apache-2.0 | |
| pipeline_tag: text-generation | |
| library_name: mlx | |
| base_model: 01-ai/Yi-Coder-9B-Chat | |
| base_model_relation: quantized | |
| quantized_by: Outlier-Ai | |
| tags: | |
| - 01-ai | |
| - 4-bit | |
| - 4bit | |
| - apple-silicon | |
| - chat | |
| - code | |
| - code-generation | |
| - coding | |
| - conversational | |
| - edge-ai | |
| - instruct | |
| - local-llm | |
| - m1 | |
| - m2 | |
| - m3 | |
| - m4 | |
| - mac | |
| - mac-mini | |
| - mac-studio | |
| - macbook-air | |
| - macbook-pro | |
| - macos | |
| - metal | |
| - mlx | |
| - mlx-community | |
| - mlx-lm | |
| - no-cloud | |
| - offline | |
| - on-device | |
| - outlier | |
| - outlier-app | |
| - private | |
| - private-ai | |
| - quantized | |
| - safetensors | |
| - text-generation | |
| - yi | |
| - yi-coder | |
| language: | |
| - en | |
| - zh | |
| widget: | |
| - example_title: Python prime check | |
| messages: | |
| - role: user | |
| content: Write a Python function that checks if a number is prime, with a short | |
| docstring and one edge case test. | |
| - example_title: SQL join | |
| messages: | |
| - role: user | |
| content: Write a SQL query that returns the top 10 users by total purchase value | |
| in the last 90 days. | |
| - example_title: Debug refactor | |
| messages: | |
| - role: user | |
| content: Refactor this for readability and flag one bug. | |
| > **Run this on your Mac with [Outlier](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=yi_coder_9b_chat_mlx_4bit)** β a free macOS app for local MLX inference. | |
| # Yi-Coder-9B-Chat (MLX 4-bit) | |
| MLX 4-bit conversion of [`01-ai/Yi-Coder-9B-Chat`](https://huggingface.co/01-ai/Yi-Coder-9B-Chat). License and base-model fields inherit from the original β see YAML frontmatter above. | |
| ## Load with mlx-lm | |
| ```bash | |
| pip install mlx-lm | |
| python -m mlx_lm.generate --model Outlier-Ai/Yi-Coder-9B-Chat-MLX-4bit --prompt "Hello" --max-tokens 256 | |
| ``` | |
| ## What is Outlier? | |
| A free macOS app that runs MLX models locally β no cloud, no API keys, no usage caps. | |
| β‘ **[outlier.host](https://outlier.host/?utm_source=hf&utm_medium=modelcard&utm_campaign=yi_coder_9b_chat_mlx_4bit)** | |
| ## Other Outlier conversions | |
| - [DeepSeek-R1-Distill-Qwen-7B (MLX 4-bit) β MLX 4-bit conversion (1,932 downloads)](https://huggingface.co/Outlier-Ai/DeepSeek-R1-Distill-Qwen-7B-MLX-4bit) | |
| - [Qwen3-Coder-30B-A3B-Instruct (MLX 4-bit) β MLX 4-bit conversion (1,598 downloads)](https://huggingface.co/Outlier-Ai/Qwen3-Coder-30B-A3B-Instruct-MLX-4bit) | |
| - [Qwen2.5-Coder-7B-Instruct (MLX 4-bit) β MLX 4-bit conversion (1,450 downloads)](https://huggingface.co/Outlier-Ai/Qwen2.5-Coder-7B-Instruct-MLX-4bit) | |
| - [Outlier-Core-27B (MLX 4-bit) β MLX 4-bit conversion (55 downloads)](https://huggingface.co/Outlier-Ai/Outlier-Core-27B-MLX-4bit) | |
| - [Outlier-Nano-4B (MLX 4-bit) β MLX 4-bit conversion (67 downloads)](https://huggingface.co/Outlier-Ai/Outlier-Nano-4B-MLX-4bit) | |
| ## License | |
| Inherits from upstream (`apache-2.0`). See base model card. | |