Instructions to use NoCanGo/WaifuGPT with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use NoCanGo/WaifuGPT with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf NoCanGo/WaifuGPT:Q8_0 # Run inference directly in the terminal: llama cli -hf NoCanGo/WaifuGPT:Q8_0
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf NoCanGo/WaifuGPT:Q8_0 # Run inference directly in the terminal: llama cli -hf NoCanGo/WaifuGPT:Q8_0
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf NoCanGo/WaifuGPT:Q8_0 # Run inference directly in the terminal: ./llama-cli -hf NoCanGo/WaifuGPT:Q8_0
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf NoCanGo/WaifuGPT:Q8_0 # Run inference directly in the terminal: ./build/bin/llama-cli -hf NoCanGo/WaifuGPT:Q8_0
Use Docker
docker model run hf.co/NoCanGo/WaifuGPT:Q8_0
- LM Studio
- Jan
- Ollama
How to use NoCanGo/WaifuGPT with Ollama:
ollama run hf.co/NoCanGo/WaifuGPT:Q8_0
- Unsloth Studio
How to use NoCanGo/WaifuGPT with Unsloth Studio:
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for NoCanGo/WaifuGPT to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for NoCanGo/WaifuGPT to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required # Open https://huggingface.co/spaces/unsloth/studio in your browser # Search for NoCanGo/WaifuGPT to start chatting
- Docker Model Runner
How to use NoCanGo/WaifuGPT with Docker Model Runner:
docker model run hf.co/NoCanGo/WaifuGPT:Q8_0
- Lemonade
How to use NoCanGo/WaifuGPT with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull NoCanGo/WaifuGPT:Q8_0
Run and chat with the model
lemonade run user.WaifuGPT-Q8_0
List all available models
lemonade list
- Atomic Chat
Install from WinGet (Windows)
winget install llama.cpp
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf NoCanGo/WaifuGPT:Q8_0# Run inference directly in the terminal:
llama cli -hf NoCanGo/WaifuGPT:Q8_0Use pre-built binary
# Download pre-built binary from:
# https://github.com/ggerganov/llama.cpp/releases# Start a local OpenAI-compatible server with a web UI:
./llama-server -hf NoCanGo/WaifuGPT:Q8_0# Run inference directly in the terminal:
./llama-cli -hf NoCanGo/WaifuGPT:Q8_0Build from source code
git clone https://github.com/ggerganov/llama.cpp.git
cd llama.cpp
cmake -B build
cmake --build build -j --target llama-server llama-cli# Start a local OpenAI-compatible server with a web UI:
./build/bin/llama-server -hf NoCanGo/WaifuGPT:Q8_0# Run inference directly in the terminal:
./build/bin/llama-cli -hf NoCanGo/WaifuGPT:Q8_0Use Docker
docker model run hf.co/NoCanGo/WaifuGPT:Q8_0A project to create a chatbot that is knowledgeable in anime since the base models are severely lacking in that department.
WaifuGPT3: Llama 3 8B WaifuGPT3.1: Llama 3.1 8B WaifuGPT3.2: Llama 3.2 3B
Version 3 has less censorship and similar performance to 3.1. Version 3.1 is newer with similar quality than 3. Version 3.2 is less accurate but gives much quicker answers, perfect for an android device.
Recommended temperature 0.5 for an accurate answer. Recommended temperature 0.7 for a creative answer.
Usage example: Who would win a fight between Sakura Haruno and Tsunade? Expected answer example: Both Sakura Haruno and Tsunade are known for their physical strength but Tsunade's experience might give her a small edge. Usage example: Answer me as Tsunade to: What do you think about Sakura Haruno? Expected answer example: Heh Sakura is a skilled kunoichi but a little too fragile if you ask me.
- Downloads last month
- 160
8-bit
Model tree for NoCanGo/WaifuGPT
Base model
meta-llama/Llama-3.1-8B
Install (macOS, Linux)
# Start a local OpenAI-compatible server with a web UI: llama serve -hf NoCanGo/WaifuGPT:Q8_0# Run inference directly in the terminal: llama cli -hf NoCanGo/WaifuGPT:Q8_0