Locutusque/hercules-v5.0
Viewer • Updated • 1.58M • 32 • 16
How to use mlx-community/Hercules-5.0-Qwen2-1.5B-4bits with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir Hercules-5.0-Qwen2-1.5B-4bits mlx-community/Hercules-5.0-Qwen2-1.5B-4bits
The Model mlx-community/Hercules-5.0-Qwen2-1.5B-4bits was converted to MLX format from M4-ai/Hercules-5.0-Qwen2-1.5B using mlx-lm version 0.14.0.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("mlx-community/Hercules-5.0-Qwen2-1.5B-4bits")
response = generate(model, tokenizer, prompt="hello", verbose=True)
Quantized