openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 7.9M • • 3.25k
Generate images from text prompts with FLUX.1 schnell
Generate speech from text using a reference voice
Transcribe audio or YouTube video into text
Generate Talking avatars from Text-to-Speech
Upscale images by 4× with a single click
Enhance and upscale images with tile‑based AI control
Transcribe audio to text instantly using WebGPU
Clone a voice and generate speech from text
Perform multiple NLP tasks like NER, QA, and summarization