--- license: apache-2.0 base_model: [bravesoftware/Ocelot-1-VL, Qwen/Qwen3-VL-4B-Instruct] library_name: mlx pipeline_tag: image-text-to-text tags: [mlx, qwen3-vl, vision, summarization, 8-bit] --- # Ocelot-1-VL MLX 8-bit Near-BF16 MLX 8-bit, group-size 64 conversion of [Ocelot-1-VL](https://huggingface.co/bravesoftware/Ocelot-1-VL), merged into its BF16 Qwen3-VL-4B-Instruct base. Effective quantization is 9.202 bits/weight because sensitive and unsupported tensors remain at higher precision. This model is specialized only for webpage summarization. Follow the strict prompt contract and limitations in the original model card. ```bash pip install 'mlx-vlm @ git+https://github.com/Blaizzy/mlx-vlm.git' python -m mlx_vlm generate --model . --prompt 'The is the text of a webpage: Page text here Summarise the content between the tags, or if no content is found use the screenshots provided, in the Brave Summary style.' --max-tokens 512 ``` Converted with MLX-VLM revision `0b1d25e334686bd36dda71b2307d186dbb3e7859`. Text generation smoke test passed.