TomLucidor commited on
Commit
8d99e25
·
verified ·
1 Parent(s): 0d4eaad

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +38 -0
README.md ADDED
@@ -0,0 +1,38 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: transformers
3
+ tags:
4
+ - compression
5
+ - expert-merging
6
+ - moe
7
+ - code
8
+ - mlx
9
+ - mlx-my-repo
10
+ license: apache-2.0
11
+ base_model: bknyaz/Qwen3-Coder-Next-REAM
12
+ ---
13
+
14
+ # TomLucidor/Qwen3-Coder-Next-REAM-mlx-4Bit
15
+
16
+ The Model [TomLucidor/Qwen3-Coder-Next-REAM-mlx-4Bit](https://huggingface.co/TomLucidor/Qwen3-Coder-Next-REAM-mlx-4Bit) was converted to MLX format from [bknyaz/Qwen3-Coder-Next-REAM](https://huggingface.co/bknyaz/Qwen3-Coder-Next-REAM) using mlx-lm version **0.29.1**.
17
+
18
+ ## Use with mlx
19
+
20
+ ```bash
21
+ pip install mlx-lm
22
+ ```
23
+
24
+ ```python
25
+ from mlx_lm import load, generate
26
+
27
+ model, tokenizer = load("TomLucidor/Qwen3-Coder-Next-REAM-mlx-4Bit")
28
+
29
+ prompt="hello"
30
+
31
+ if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
32
+ messages = [{"role": "user", "content": prompt}]
33
+ prompt = tokenizer.apply_chat_template(
34
+ messages, tokenize=False, add_generation_prompt=True
35
+ )
36
+
37
+ response = generate(model, tokenizer, prompt=prompt, verbose=True)
38
+ ```