File size: 1,083 Bytes
7e15704
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
---
license: apache-2.0
base_model: [bravesoftware/Ocelot-1-VL, Qwen/Qwen3-VL-4B-Instruct]
library_name: mlx
pipeline_tag: image-text-to-text
tags: [mlx, qwen3-vl, vision, summarization, 8-bit]
---

# Ocelot-1-VL MLX 8-bit

Near-BF16 MLX 8-bit, group-size 64 conversion of [Ocelot-1-VL](https://huggingface.co/bravesoftware/Ocelot-1-VL), merged into its BF16 Qwen3-VL-4B-Instruct base. Effective quantization is 9.202 bits/weight because sensitive and unsupported tensors remain at higher precision.

This model is specialized only for webpage summarization. Follow the strict prompt contract and limitations in the original model card.

```bash
pip install 'mlx-vlm @ git+https://github.com/Blaizzy/mlx-vlm.git'
python -m mlx_vlm generate --model . --prompt 'The is the text of a webpage: <page>Page text here</page> Summarise the content between the <page> tags, or if no content is found use the screenshots provided, in the Brave Summary style.' --max-tokens 512
```

Converted with MLX-VLM revision `0b1d25e334686bd36dda71b2307d186dbb3e7859`. Text generation smoke test passed.