--- license: apache-2.0 --- --- base_model: unsloth/Meta-Llama-3.1-8B-bnb-4bit library_name: transformers tags: - unsloth - llama-3 - kubernetes - devops - gguf - text-generation license: apache-2.0 language: - en --- # Kubernetes Expert Llama-3.1-8B (GGUF) This model is a fine-tuned version of **Llama-3.1-8B**, specifically trained to be a **Kubernetes Expert**. It was trained using [Unsloth](https://github.com/unslothai/unsloth) and LoRA on a high-quality StackOverflow Kubernetes dataset. ## 🚀 Model Features - **Format**: GGUF (Quantized to `Q4_K_M`) - **Use Case**: Answering technical K8s questions, debugging pods (`CrashLoopBackOff`), and generating YAML configurations. - **Performance**: 2x faster inference speed compared to the baseline model with significantly better domain knowledge. ## 📦 How to Use with Ollama 1. **Download the Model** Download `k8s-expert-rescue.Q4_K_M.gguf` from the Files tab. 2. **Create a Modelfile** Create a file named `Modelfile` with the following content: ```dockerfile FROM ./k8s-expert-rescue.Q4_K_M.gguf TEMPLATE """Below is an instruction that describes a task, paired with an input that provides further context. Write a response that appropriately completes the request. ### Instruction: You are a Kubernetes expert. Provide a technical solution to the following problem. ### Input: {{ .Prompt }} ### Response: """ PARAMETER temperature 0.6 PARAMETER num_ctx 4096 PARAMETER stop "<|end_of_text|>" PARAMETER stop "### Instruction:" PARAMETER stop "### Input:"