Running on Zero Agents Featured 5.14k FLUX.1 [Schnell] π 5.14k Generate images from text prompts with FLUX.1-schnell
Running Agents Featured 1.14k OpenVoice π€ 1.14k Generate speech in a cloned voice from a short audio clip
Running on CPU Upgrade Featured 976 TTS Arena V2 π£ 976 Compare two text-to-speech voices and vote for the better
Runtime error Agents 408 HierSpeech++ (Zero-shot TTS) β‘ 408 Generate high-quality speech from text using a prompt audio
Running on Zero Agents Featured 399 Playground V2 π 399 Generate images from text prompts with customizable options
playgroundai/playground-v2-1024px-aesthetic Text-to-Image β’ 3B β’ Updated Feb 23, 2024 β’ 329 β’ 560
pyannote/speaker-diarization-3.1 Automatic Speech Recognition β’ Updated May 10, 2024 β’ 7.27M β’ 3.94k
Running on Zero Agents Featured 734 StyleTTS 2 π£ 734 Efficient, fast, and natural text to speech with StyleTTS 2!
stabilityai/stable-video-diffusion-img2vid-xt Image-to-Video β’ 2B β’ Updated Jul 10, 2024 β’ 241k β’ 3.43k
Running Agents 189 Gradio Lipsync Wav2lip π 189 Generate lipβsynced video from a face image and audio
pyannote/speaker-diarization Automatic Speech Recognition β’ Updated May 10, 2024 β’ 340k β’ 1.34k