Configuration Parsing Warning:In config.json: "quantization_config.modules_to_not_convert" must be an array

Optimized to run on the Nvidia Jetson Orin Nano.

Downloads last month
12
Safetensors
Model size
3B params
Tensor type
I32
·
F16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for jsbaicenter/Llama-3.2-3b-Instruct-AWQ-4bit-GEMM

Quantized
(508)
this model

Collection including jsbaicenter/Llama-3.2-3b-Instruct-AWQ-4bit-GEMM