Audio-to-Audio
Transformers
Safetensors
multilingual
speecht5
philippines
philippine-languages
voice-conversion
speech-to-speech
bcl
ceb
eng
fil
hil
ilo
pag
pam
tsg
war
Instructions to use sapinsapin/speecht5_vc-pld with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use sapinsapin/speecht5_vc-pld with Transformers:
# Load model directly from transformers import AutoProcessor, SpeechT5ForSpeechToSpeechWithLoss processor = AutoProcessor.from_pretrained("sapinsapin/speecht5_vc-pld") model = SpeechT5ForSpeechToSpeechWithLoss.from_pretrained("sapinsapin/speecht5_vc-pld", device_map="auto") - Notebooks
- Google Colab
- Kaggle
| { | |
| "feature_extractor": { | |
| "do_normalize": false, | |
| "feature_extractor_type": "SpeechT5FeatureExtractor", | |
| "feature_size": 1, | |
| "fmax": 7600, | |
| "fmin": 80, | |
| "frame_signal_scale": 1.0, | |
| "hop_length": 16, | |
| "mel_floor": 1e-10, | |
| "num_mel_bins": 80, | |
| "padding_side": "right", | |
| "padding_value": 0.0, | |
| "return_attention_mask": true, | |
| "sampling_rate": 16000, | |
| "win_function": "hann_window", | |
| "win_length": 64 | |
| }, | |
| "processor_class": "SpeechT5Processor" | |
| } | |