Audio-to-Audio
Transformers
Safetensors
multilingual
speecht5
philippines
philippine-languages
voice-conversion
speech-to-speech
bcl
ceb
eng
fil
hil
ilo
pag
pam
tsg
war
Instructions to use sapinsapin/speecht5_vc-pld with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use sapinsapin/speecht5_vc-pld with Transformers:
# Load model directly from transformers import AutoProcessor, SpeechT5ForSpeechToSpeechWithLoss processor = AutoProcessor.from_pretrained("sapinsapin/speecht5_vc-pld") model = SpeechT5ForSpeechToSpeechWithLoss.from_pretrained("sapinsapin/speecht5_vc-pld", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 492 Bytes
ed0b31b | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 | {
"feature_extractor": {
"do_normalize": false,
"feature_extractor_type": "SpeechT5FeatureExtractor",
"feature_size": 1,
"fmax": 7600,
"fmin": 80,
"frame_signal_scale": 1.0,
"hop_length": 16,
"mel_floor": 1e-10,
"num_mel_bins": 80,
"padding_side": "right",
"padding_value": 0.0,
"return_attention_mask": true,
"sampling_rate": 16000,
"win_function": "hann_window",
"win_length": 64
},
"processor_class": "SpeechT5Processor"
}
|