pozdgpt / README.md
sodeeplearning's picture
Update README.md
9ee2b74 verified
|
Raw
History Blame Contribute Delete
1.13 kB
---
license: mit
language:
- ru
---
# PozdGPT
Peak of humanity technologies. New breath in neuroslop world.
![PozdGPT](./pozdgpt.jpg)
## Usage
Via llama.cpp + GGUF
```python
# !pip install llama-cpp-python
from llama_cpp import Llama
llm = Llama.from_pretrained(
repo_id="sodeeplearning/pozdgpt",
filename="PozdGPT-Q4_K_M.gguf", # Or Q6_K, Q8_0, f16
)
```
## AWQ + vLLM
To launch 4bit AWQ version you need to download this
[folder](https://huggingface.co/sodeeplearning/pozdgpt/tree/main/PozdGPT-awq-4bit)
and launch your vLLM server:
```bash
# !pip install vllm
vllm serve ./PozdGPT-awq-4bit \
--served-model-name pozdgpt \
--quantization compressed-tensors \
--max-model-len 8192 \
--gpu-memory-utilization 0.88 \
--max-num-seqs 6 \
--kv-cache-dtype fp8 \
--enable-prefix-caching \
--api-key key \
--port 8148
```
## Test via telegram bot
You can test this bot in official [telegram bot](https://t.me/pozdgpt_bot)
# Contacts
- [Github](https://github.com/sodeeplearning/PozdGPT)
- [Team Telegram](https://t.me/Notfag)
- [Telegram bot](https://t.me/pozdgpt_bot)
- Email: vitaliy.petreev@gmail.com