Lizzy-7B-8bits-mlx

Introduce

Quant by mlx-lm 0.31.3 with new pipeline for lizzy model_type

ref

Usage

mlx_lm.generate mlx_lm.generate --model /path/to/model --prompt "What is different between CPU and GPU?" --max-tokens 2048 --trust-remote-code
# or
mlx_lm.chat --model /path/to/model --max-tokens 2048 --trust-remote-code
Downloads last month
19
Safetensors
Model size
2B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AlphaOxO/Lizzy-7B-8bits-MLX

Quantized
(3)
this model

Collection including AlphaOxO/Lizzy-7B-8bits-MLX