Llama 3.2 3B Instruct โ€” Solus v1

The larger of Meta's two small Llama 3.2 instruction-tuned models. At roughly 3.2B parameters it is a genuine general-purpose assistant that still fits comfortably on a laptop, and it is a noticeable step up from the 1B on instruction following, summarisation, and multi-turn chat.

Officially supports eight languages. A good default when you want everyday chat and writing quality without a discrete GPU.

Specifications

Parameters 3B
Quantization Q4_K_M
File size 1.88 GB
Minimum RAM 5.00 GB
Minimum VRAM not required
Context length 16,384 tokens
SHA-256 6c1a2b41161032677be168d354123594c0e6e67d2b9227c84f296ad037c728ff

Single file: Llama-3.2-3B-Instruct-Q4_K_M.gguf

Quantization

Quantization performed at the Faculty of Engineering, McMaster University.

The GGUF conversion this build is derived from was produced by bartowski, and the weights here are a byte-for-byte copy of that file โ€” the SHA-256 above matches the upstream artifact.

Provenance

Usage

llama-cli -m Llama-3.2-3B-Instruct-Q4_K_M.gguf -cnv

License

Built with Llama.

Llama 3.2 is licensed under the Llama 3.2 Community License, Copyright (c) Meta Platforms, Inc. All Rights Reserved. Your use of this model is governed by that license and by the Llama 3.2 Acceptable Use Policy.

Downloads last month
11
GGUF
Model size
3B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Fazmin/solus_v1_llama-3.2-3b-instruct-q4

Quantized
(494)
this model