| --- |
| license: cc-by-nc-sa-4.0 |
| library_name: vokra |
| tags: |
| - vokra |
| - gguf |
| --- |
| |
| # audioldm2 (Vokra GGUF) |
|
|
| Converted to the Vokra GGUF format for [Vokra](https://github.com/ayutaz/vokra), a zero-dependency speech-AI inference runtime. |
|
|
| **This is a conversion, not a new model.** The weights are the upstream ones; Vokra re-packages them so its runtime can memory-map them directly. Credit for the model belongs upstream β see *Source* below. |
|
|
| ## Files |
|
|
| | File | Size | SHA-256 | |
| |---|---|---| |
| | `audioldm2.gguf` | 4266.4 MB | `d02dd97944b05bd27ea17eb0c31034c27992c0f60e642cffef28047a436cb5f1` | |
|
|
| ## Usage |
|
|
| ```bash |
| # Download (any HTTP client works β the file is a plain GGUF) |
| curl -L -o audioldm2.gguf \ |
| https://huggingface.co/vokra/audioldm2/resolve/main/audioldm2.gguf |
| ``` |
|
|
| ```bash |
| vokra-cli run --model audioldm2.gguf --input input.wav |
| ``` |
|
|
| ## Provenance |
|
|
| | Field | Value | |
| |---|---| |
| | Architecture | `audioldm2` | |
| | Tensors | 2827 | |
| | Upstream source | cvssp/audioldm2 (Liu et al. 2024 arXiv:2308.05734 text-to-audio LDM, cc-by-nc-sa-4.0) | |
| | Upstream licence | `cc-by-nc-sa-4.0` | |
| | Licence class | `non-commercial-share-alike` | |
| | Registry model id | `audioldm2` | |
| | Vokra GGUF schema | 1 | |
| | Converted by | vokra-core 0.1.0-alpha.0 | |
|
|
| Every row above is read out of this file's own `vokra.*` metadata, so the card cannot claim something the artifact does not carry. |
|
|
| ## Licence |
|
|
| The weights are distributed under **cc-by-nc-sa-4.0**, unchanged from upstream. Conversion does not alter the licence, and your obligations run to the upstream author. |
|
|
| ### β οΈ Share-alike / copyleft β this licence travels with the file |
|
|
| This weight is **cc-by-nc-sa-4.0**, and that licence is *not* discharged by attribution alone. It attaches to derivatives. |
|
|
| - This GGUF is a **conversion of an upstream weight**, so it is itself **cc-by-nc-sa-4.0** β not Apache-2.0, and not covered by Vokra's own licence. |
| - Anything you derive from it (a fine-tune, a re-quantisation, a further format conversion) carries the same licence. |
| - Vokra's runtime is Apache-2.0. **Loading this weight does not change that**, because these licences restrict the terms of redistribution, not use. Shipping the *weight* onward is what carries the obligation. |
|
|
| ### β Non-commercial β you may not use this weight commercially |
|
|
| The upstream weight is **cc-by-nc-sa-4.0**. It is republished here so the model can be evaluated and used for research, and the licence is unchanged by conversion. |
|
|
| - **Do not use this in a commercial product or service.** That restriction is upstream's, not Vokra's, and Vokra cannot waive it. |
| - Vokra's engine is Apache-2.0 and imposes no such limit β the limit is on **this weight**. Other models in this organisation are permissively licensed; check each one's card. |
| - Vokra's runtime refuses to load a non-commercial weight unless an explicit research flag is set, so this restriction is enforced at load time rather than left to the reader. |
|
|
| ## Verifying this file |
|
|
| ```bash |
| shasum -a 256 audioldm2.gguf |
| # expect: d02dd97944b05bd27ea17eb0c31034c27992c0f60e642cffef28047a436cb5f1 |
| ``` |
|
|