How to use from
llama.cpp
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf concedo/Beepo-22B-GGUF:
# Run inference directly in the terminal:
llama cli -hf concedo/Beepo-22B-GGUF:
Install from WinGet (Windows)
winget install llama.cpp
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf concedo/Beepo-22B-GGUF:
# Run inference directly in the terminal:
llama cli -hf concedo/Beepo-22B-GGUF:
Use pre-built binary
# Download pre-built binary from:
# https://github.com/ggerganov/llama.cpp/releases
# Start a local OpenAI-compatible server with a web UI:
./llama-server -hf concedo/Beepo-22B-GGUF:
# Run inference directly in the terminal:
./llama-cli -hf concedo/Beepo-22B-GGUF:
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git
cd llama.cpp
cmake -B build
cmake --build build -j --target llama-server llama-cli
# Start a local OpenAI-compatible server with a web UI:
./build/bin/llama-server -hf concedo/Beepo-22B-GGUF:
# Run inference directly in the terminal:
./build/bin/llama-cli -hf concedo/Beepo-22B-GGUF:
Use Docker
docker model run hf.co/concedo/Beepo-22B-GGUF:
Quick Links

Beepo-22B-GGUF

This is the GGUF quantization of https://huggingface.co/concedo/Beepo-22B, which was originally finetuned on top of the https://huggingface.co/mistralai/Mistral-Small-Instruct-2409 model.

You can use KoboldCpp to run this model.

image/png

Key Features:

  • Retains Intelligence - LR was kept low and dataset heavily pruned to avoid losing too much of the original model's intelligence.
  • Instruct prompt format supports Alpaca - Honestly, I don't know why more models don't use it. If you are an Alpaca format lover like me, this should help. The original Mistral instruct format can still be used, but is not recommended.
  • Instruct Decensoring Applied - You should not need a jailbreak for a model to obey the user. The model should always do what you tell it to. No need for weird "Sure, I will" or kitten-murdering-threat tricks. No abliteration was done, only finetuning. This model is not evil. It does not judge or moralize. Like a good tool, it simply obeys.

Prompt template: Alpaca

### Instruction:
{prompt}

### Response:

Please leave any feedback or issues that you may have.

Downloads last month
827
GGUF
Model size
22B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

2-bit

4-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support