GGUF
k8s-llama-expert / README.md
FinysterLin's picture
Update README.md
63bd3ee verified
|
Raw
History Blame Contribute Delete
1.58 kB
---
license: apache-2.0
---
---
base_model: unsloth/Meta-Llama-3.1-8B-bnb-4bit
library_name: transformers
tags:
- unsloth
- llama-3
- kubernetes
- devops
- gguf
- text-generation
license: apache-2.0
language:
- en
---
# Kubernetes Expert Llama-3.1-8B (GGUF)
This model is a fine-tuned version of **Llama-3.1-8B**, specifically trained to be a **Kubernetes Expert**. It was trained using [Unsloth](https://github.com/unslothai/unsloth) and LoRA on a high-quality StackOverflow Kubernetes dataset.
## 🚀 Model Features
- **Format**: GGUF (Quantized to `Q4_K_M`)
- **Use Case**: Answering technical K8s questions, debugging pods (`CrashLoopBackOff`), and generating YAML configurations.
- **Performance**: 2x faster inference speed compared to the baseline model with significantly better domain knowledge.
## 📦 How to Use with Ollama
1. **Download the Model**
Download `k8s-expert-rescue.Q4_K_M.gguf` from the Files tab.
2. **Create a Modelfile**
Create a file named `Modelfile` with the following content:
```dockerfile
FROM ./k8s-expert-rescue.Q4_K_M.gguf
TEMPLATE """Below is an instruction that describes a task, paired with an input that provides further context. Write a response that appropriately completes the request.
### Instruction:
You are a Kubernetes expert. Provide a technical solution to the following problem.
### Input:
{{ .Prompt }}
### Response:
"""
PARAMETER temperature 0.6
PARAMETER num_ctx 4096
PARAMETER stop "<|end_of_text|>"
PARAMETER stop "### Instruction:"
PARAMETER stop "### Input:"