Text Generation
Safetensors
English
llama
dpo
preference-alignment
fine-tuned
unsloth
lora
nlp
deep-learning
gordon-ramsay
conversational
Eval Results (legacy)
Instructions to use antonisbast/Llama-3.2-3B-Gordon-Ramsay-DPO with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Local Apps Settings
- Unsloth Desktop
Update README.md
Browse files
README.md
CHANGED
|
@@ -170,7 +170,7 @@ This model was also integrated into a custom **Retrieval-Augmented Generation (R
|
|
| 170 |
## Acknowledgments
|
| 171 |
|
| 172 |
- **Course:** AIDL_B_CS01 — Natural Language Processing with Deep Learning, University of West Attica
|
| 173 |
-
- **Instructor:**
|
| 174 |
- **Base Model:** [Meta Llama 3.2](https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct) under the Llama 3.2 Community License
|
| 175 |
- **Training Framework:** [Unsloth](https://github.com/unslothai/unsloth) for efficient LoRA fine-tuning---
|
| 176 |
|
|
|
|
| 170 |
## Acknowledgments
|
| 171 |
|
| 172 |
- **Course:** AIDL_B_CS01 — Natural Language Processing with Deep Learning, University of West Attica
|
| 173 |
+
- **Instructor:** Panagiotis Kasnesis
|
| 174 |
- **Base Model:** [Meta Llama 3.2](https://huggingface.co/meta-llama/Llama-3.2-3B-Instruct) under the Llama 3.2 Community License
|
| 175 |
- **Training Framework:** [Unsloth](https://github.com/unslothai/unsloth) for efficient LoRA fine-tuning---
|
| 176 |
|