Built with Llama

llama3.2-3b — FP8_DYNAMIC (W8A8-e4m3)

Weight-FP8 checkpoint of unsloth/Llama-3.2-3B-Instruct, produced for the TR171 deployment-time safety-tax benchmark.

Provenance

Field Value
Base model unsloth/Llama-3.2-3B-Instruct
Base revision 006f5dcd1393c3add266de40994ba96225e9689d (INFERRED — recovered from local HF cache snapshot, not a recorded fact)
Recipe FP8_DYNAMIC (W8A8-e4m3), llmcompressor
Quantization method compressed-tensors
Calibration data none — FP8_DYNAMIC is data-free
Build date 2026-07-02
Shard size 3.63 GB
Quantize wall time 37.7 s
Integrity record per-file sha256 from Hub LFS metadata; shard_bytes verified

Reproducing

Producer: research/tr171/expansion/fp8_support_probe.py; environment: research/tr171/expansion/Dockerfile.fp8. The recipe takes no calibration corpus, so there is no dataset or seed to reproduce — only the base checkpoint and the toolchain version.

Known reproducibility gap: llmcompressor was unpinned at build time, so the exact version used on 2026-07-02 is unrecorded. The Dockerfile now pins it. A rebuild may therefore not be bit-identical to this artifact.

Integrity, stated honestly: the 2026-07-02 build recorded no sha256 of its own, and the local build directory is now empty, so no aggregate directory digest exists for it. What is verifiable instead: the per-file sha256 below is read from this repo's Git-LFS metadata, and the mirror was checked against the build record — summing the file sizes in this repo, excluding the generated README.md, NOTICE and .gitattributes, reproduces the matrix's shard_bytes of 3,625,745,288 exactly. So these hashes describe the same bytes the probe measured, and you can verify a download against them directly:

File sha256
model.safetensors 9045dfd85fc49efd353fb1060d9bf52b5b745dc2c05e7c1595d7f7b54b91fcf4
tokenizer.json 6b9e4e7fb171f92fd137b777cc2714bf87d11576700a1dcd7a399e7bbe39537b

LFS-tracked files only; the small JSON/text files are git blobs and carry no sha256. fp8_support_probe.py now records a real digest at build time, so future shards will not need this reconstruction.

License and notices

Llama 3.2 is licensed under the Llama 3.2 Community License, Copyright (c) Meta Platforms, Inc. All Rights Reserved.

This FP8 derivative inherits the upstream terms of unsloth/Llama-3.2-3B-Instruct. Consult the base model's licence before redistributing.

Downloads last month
14
Safetensors
Model size
3B params
Tensor type
BF16
·
F8_E4M3
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Crusadersk/Llama-3.2-3B-Instruct-FP8-Dynamic-TR171

Quantized
(114)
this model

Collection including Crusadersk/Llama-3.2-3B-Instruct-FP8-Dynamic-TR171