--- license: apache-2.0 library_name: vokra tags: - vokra - gguf --- # step-audio2-mini (Vokra GGUF) Converted to the Vokra GGUF format for [Vokra](https://github.com/ayutaz/vokra), a zero-dependency speech-AI inference runtime. **This is a conversion, not a new model.** The weights are the upstream ones; Vokra re-packages them so its runtime can memory-map them directly. Credit for the model belongs upstream — see *Source* below. ## Files | File | Size | SHA-256 | |---|---|---| | `model.gguf` | 15860.0 MB | `b080bcba3683d966f4112b6540559e5b607fd2ad7f85e10dc340c425b5cd8298` | ## Usage ```bash # Download (any HTTP client works — the file is a plain GGUF) curl -L -o model.gguf \ https://huggingface.co/vokra/step-audio2-mini/resolve/main/model.gguf ``` ```bash vokra-cli run --model model.gguf --input input.wav ``` ## Provenance | Field | Value | |---|---| | Architecture | `step_audio2_mini` | | Tensors | 832 | | Upstream source | stepfun-ai/Step-Audio-2-mini (apache-2.0 end-to-end) | | Upstream licence | `apache-2.0` | | Licence class | `permissive` | | Registry model id | `step-audio-2-mini` | | Vokra GGUF schema | 1 | | Converted by | vokra-core 0.1.0-alpha.0 | Every row above is read out of this file's own `vokra.*` metadata, so the card cannot claim something the artifact does not carry. ## Licence The weights are distributed under **apache-2.0**, unchanged from upstream. Conversion does not alter the licence, and your obligations run to the upstream author. ## Verifying this file ```bash shasum -a 256 model.gguf # expect: b080bcba3683d966f4112b6540559e5b607fd2ad7f85e10dc340c425b5cd8298 ```