Dataset
#3
by yukiarimo - opened
How many total audio files (and what’s the hour count across them) have used to train the model? Is it like LJSpeech-sized?
And is voice real or synthetic (TTS distillation)?
This comment has been hidden
How many total audio files (and what’s the hour count across them) have used to train the model? Is it like LJSpeech-sized?
And is voice real or synthetic (TTS distillation)?