Sharanya186
/

TextToSpeech

speech-synthesis

Model card Files Files and versions

Sharanya186 commited on Apr 21, 2024

Commit

0189ccf

·

verified ·

1 Parent(s): 5dc6008

Update README.md

Files changed (1) hide show

README.md +26 -0

README.md CHANGED Viewed

@@ -17,3 +17,29 @@ datasets:
 This repository provides all the necessary tools for Text-to-Speech (TTS)  with SpeechBrain using a [Transformer](https://arxiv.org/pdf/1809.08895.pdf) pretrained on [LJSpeech](https://keithito.com/LJ-Speech-Dataset/).
 The pre-trained model takes in text input and produces a spectrogram in output. One can get the final waveform by applying a vocoder (e.g., HiFIGAN) on top of the generated spectrogram.

 This repository provides all the necessary tools for Text-to-Speech (TTS)  with SpeechBrain using a [Transformer](https://arxiv.org/pdf/1809.08895.pdf) pretrained on [LJSpeech](https://keithito.com/LJ-Speech-Dataset/).
 The pre-trained model takes in text input and produces a spectrogram in output. One can get the final waveform by applying a vocoder (e.g., HiFIGAN) on top of the generated spectrogram.
+### Perform Text-to-Speech (TTS)
+```python
+import torchaudio
+from speechbrain.inference.vocoders import HIFIGAN
+texts = ["This is a example for synthesis."]
+#initializing my model
+my_tts_model = TextToSpeech.from_hparams(source="/content/")
+#initializing vocoder(Hifigan) model
+hifi_gan = HIFIGAN.from_hparams(source="speechbrain/tts-hifigan-ljspeech", savedir="tmpdir_vocoder")
+# Running the TTS
+mel_output = my_tts_model.encode_text(texts)
+# Running Vocoder (spectrogram-to-waveform)
+waveforms = hifi_gan.decode_batch(mel_output)
+# Save the waverform
+torchaudio.save('example_TTS.wav',waveforms.squeeze(1), 22050)
+```