How to use from
llama.cpp
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf NoCanGo/WaifuGPT:Q8_0
# Run inference directly in the terminal:
llama cli -hf NoCanGo/WaifuGPT:Q8_0
Install from WinGet (Windows)
winget install llama.cpp
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf NoCanGo/WaifuGPT:Q8_0
# Run inference directly in the terminal:
llama cli -hf NoCanGo/WaifuGPT:Q8_0
Use pre-built binary
# Download pre-built binary from:
# https://github.com/ggerganov/llama.cpp/releases
# Start a local OpenAI-compatible server with a web UI:
./llama-server -hf NoCanGo/WaifuGPT:Q8_0
# Run inference directly in the terminal:
./llama-cli -hf NoCanGo/WaifuGPT:Q8_0
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git
cd llama.cpp
cmake -B build
cmake --build build -j --target llama-server llama-cli
# Start a local OpenAI-compatible server with a web UI:
./build/bin/llama-server -hf NoCanGo/WaifuGPT:Q8_0
# Run inference directly in the terminal:
./build/bin/llama-cli -hf NoCanGo/WaifuGPT:Q8_0
Use Docker
docker model run hf.co/NoCanGo/WaifuGPT:Q8_0
Quick Links

A project to create a chatbot that is knowledgeable in anime since the base models are severely lacking in that department.

WaifuGPT3: Llama 3 8B WaifuGPT3.1: Llama 3.1 8B WaifuGPT3.2: Llama 3.2 3B

Version 3 has less censorship and similar performance to 3.1. Version 3.1 is newer with similar quality than 3. Version 3.2 is less accurate but gives much quicker answers, perfect for an android device.

Recommended temperature 0.5 for an accurate answer. Recommended temperature 0.7 for a creative answer.

Usage example: Who would win a fight between Sakura Haruno and Tsunade? Expected answer example: Both Sakura Haruno and Tsunade are known for their physical strength but Tsunade's experience might give her a small edge. Usage example: Answer me as Tsunade to: What do you think about Sakura Haruno? Expected answer example: Heh Sakura is a skilled kunoichi but a little too fragile if you ask me.

Downloads last month
160
GGUF
Model size
8B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for NoCanGo/WaifuGPT

Dataset used to train NoCanGo/WaifuGPT