google/gemma-4-E2B-it, UQFF quantization

Run with mistral.rs. Documentation: UQFF docs.

  1. Flexible 🌀: Multiple quantization formats in one file format with one framework to run them all.
  2. Reliable 🔒: Compatibility ensured with embedded and checked semantic versioning information from day 1.
  3. Easy 🤗: Download UQFF models easily and quickly from Hugging Face, or use a local file.
  4. Customizable 🛠️: Make and publish your own UQFF files in minutes.

Install

Install mistral.rs (full guide):

Linux/macOS:

curl --proto '=https' --tlsv1.2 -sSf https://raw.githubusercontent.com/EricLBuehler/mistral.rs/master/install.sh | sh

Windows (PowerShell):

irm https://raw.githubusercontent.com/EricLBuehler/mistral.rs/master/install.ps1 | iex

Examples

Note: AFQ variants are optimized for Apple Silicon / Metal.

Quantization Command
AFQ4 mistralrs run -m nchapman/gemma-4-E2B-it-UQFF multimodal-plain --from-uqff afq4-0.uqff
Q4K mistralrs run -m nchapman/gemma-4-E2B-it-UQFF multimodal-plain --from-uqff q4k-0.uqff
Downloads last month
4
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for nchapman/gemma-4-E2B-it-UQFF

Quantized
(308)
this model