Automatic speech recognition (a.k.a. speech-to-text), speech synthesis (text-to-speech, voice conversion), machine translation, simultaneous speech-to-speech translation, speech/text corpus