Hybrid Transformers for Music Source Separation
Paper • 2211.08553 • Published • 1
Converted to the Vokra GGUF format for Vokra, a zero-dependency speech-AI inference runtime.
This is a conversion, not a new model. The weights are the upstream ones; Vokra re-packages them so its runtime can memory-map them directly. Credit for the model belongs upstream — see Source below.
| File | Size | SHA-256 |
|---|---|---|
htdemucs-multi.gguf |
320.5 MB | 41c3d8ef26c94cddb1df32ea2d131bca2b0ce1d951b7fac4dc4253b835ab89f2 |
# Download (any HTTP client works — the file is a plain GGUF)
curl -L -o htdemucs-multi.gguf \
https://huggingface.co/vokra/htdemucs-multi/resolve/main/htdemucs-multi.gguf
vokra-cli run --model htdemucs-multi.gguf --input input.wav
| Field | Value |
|---|---|
| Architecture | htdemucs_multi |
| Tensors | 2132 |
| Upstream source | facebook/htdemucs family (Hybrid Transformer Demucs music source separation, 4-source htdemucs_ft or 6-source htdemucs_6s variant, Rouard et al. 2023 arXiv:2211.08553, MIT) |
| Upstream licence | mit |
| Licence class | permissive |
| Registry model id | htdemucs_multi |
| Vokra GGUF schema | 1 |
| Converted by | vokra-core 0.1.0-alpha.0 |
Every row above is read out of this file's own vokra.* metadata, so the card cannot claim something the artifact does not carry.
The weights are distributed under mit, unchanged from upstream. Conversion does not alter the licence, and your obligations run to the upstream author.
shasum -a 256 htdemucs-multi.gguf
# expect: 41c3d8ef26c94cddb1df32ea2d131bca2b0ce1d951b7fac4dc4253b835ab89f2
We're not able to determine the quantization variants.