---
license: apache-2.0
base_model: [bravesoftware/Ocelot-1-VL, Qwen/Qwen3-VL-4B-Instruct]
library_name: mlx
pipeline_tag: image-text-to-text
tags: [mlx, qwen3-vl, vision, summarization, 8-bit]
---
# Ocelot-1-VL MLX 8-bit
Near-BF16 MLX 8-bit, group-size 64 conversion of [Ocelot-1-VL](https://huggingface.co/bravesoftware/Ocelot-1-VL), merged into its BF16 Qwen3-VL-4B-Instruct base. Effective quantization is 9.202 bits/weight because sensitive and unsupported tensors remain at higher precision.
This model is specialized only for webpage summarization. Follow the strict prompt contract and limitations in the original model card.
```bash
pip install 'mlx-vlm @ git+https://github.com/Blaizzy/mlx-vlm.git'
python -m mlx_vlm generate --model . --prompt 'The is the text of a webpage: Page text here Summarise the content between the tags, or if no content is found use the screenshots provided, in the Brave Summary style.' --max-tokens 512
```
Converted with MLX-VLM revision `0b1d25e334686bd36dda71b2307d186dbb3e7859`. Text generation smoke test passed.