alst10's picture
Upload README.md with huggingface_hub
cd5bce8 verified
|
Raw History Blame Contribute Delete
1.38 kB
---
license: cc-by-4.0
base_model: alst10/ReverseLlama
tags:
- gguf
- rocm
- rocmfpx
- amd
- llama.cpp
---
# ReverseLlama-ROCMFP4_FAST - ROCmFPX Quantized
This repository contains AMD ROCmFPX quantized GGUF weights converted from [alst10/ReverseLlama](https://huggingface.co/alst10/ReverseLlama) using [charlie12345/ROCmFPX](https://github.com/charlie12345/ROCmFPX).
## Available Files
| File | Size | Format |
| --- | --- | --- |
| `ReverseLlama-Q4_0_ROCMFP4_FAST.gguf` | 3.98 GB | Q4_0_ROCMFP4_FAST |
## Usage Instructions with ROCmFPX llama.cpp
```bash
git clone --depth 1 [https://github.com/charlie12345/ROCmFPX.git](https://github.com/charlie12345/ROCmFPX.git)
cd ROCmFPX
cmake -B build -G Ninja -DCMAKE_BUILD_TYPE=Release -DGGML_CUDA=OFF -DGGML_NATIVE=ON
cmake --build build --target llama-cli -j
./build/bin/llama-cli -m ReverseLlama-Q4_0_ROCMFP4_FAST.gguf -p "You are a helpful assistant. Hello!" -n 128
```
## License
This model is distributed under the [Creative Commons Attribution 4.0 International (CC BY 4.0)](https://creativecommons.org/licenses/by/4.0/) license.