wonone commited on
Commit
8b83675
·
verified ·
1 Parent(s): b65a216

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +0 -8
README.md CHANGED
@@ -71,14 +71,6 @@ In practice, MTP speedups depend on draft acceptance rate. The grafted head tend
71
 
72
  Use oMLX and enable Native MTP.
73
 
74
- Suggested initial settings:
75
-
76
- * Native MTP: enabled
77
- * Max Draft Tokens: 2
78
- * Min Draft Tokens: 1
79
- * Temperature: 0 for benchmarking
80
- * Use the same prompt, context length, and max tokens when comparing against non-MTP variants
81
-
82
  ## Notes
83
 
84
  This is not an OptiQ oQ8 sidecar model. The model uses a native MLX-VLM layout with `vision_tower.*` weights included in the model files.
 
71
 
72
  Use oMLX and enable Native MTP.
73
 
 
 
 
 
 
 
 
 
74
  ## Notes
75
 
76
  This is not an OptiQ oQ8 sidecar model. The model uses a native MLX-VLM layout with `vision_tower.*` weights included in the model files.