Any-to-Any
ESPnet
PyTorch
English
audio
multimodal
speech-language-model
audio-understanding
audio-generation
text-to-audio
Instructions to use espnet/bagpiper-sft with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- ESPnet
How to use espnet/bagpiper-sft with ESPnet:
unknown model type (must be text-to-speech or automatic-speech-recognition)
- Notebooks
- Google Colab
- Kaggle
| dtype: bfloat16 | |
| num_hypo: 1 | |
| enforce_modality: ["text"] | |
| audio: | |
| temperature: 0.8 | |
| topk: 20 | |
| cfg: 3 | |
| max_step: 2048 | |
| min_step: 50 | |
| text: | |
| temperature: 0.6 | |
| topk: 20 | |
| cfg: 1 | |
| max_step: 2048 | |
| min_step: 1 | |