Usage

This model is used local with ollama.

  1. Download the GGUF file and the Modelfile.
  2. Create a new ollama model with ollama create <model_name> -f Modelfile.
  3. Start the model with ollama run <model_file>.

In the Modelfile you can adjust parameters like temperature, change the template or the system prompt.

Downloads last month
2
GGUF
Model size
8B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for wijan/NLP-prototype-Meta-Llama-3_1-8B-Instruct-bnb-4bit

Quantized
(5)
this model

Collection including wijan/NLP-prototype-Meta-Llama-3_1-8B-Instruct-bnb-4bit