Running on Zero Agents 709 MiniMax H3 Turbo LoRA 🎬 709 Video generation with a synchronized soundtrack
ibm-granite/granite-speech-5.0-470m-turboctc Automatic Speech Recognition • 0.5B • Updated 6 days ago • 51.1k • 69
firdhokk/speech-emotion-recognition-with-facebook-wav2vec2-large-xlsr-53 Audio Classification • 0.3B • Updated Dec 15, 2024 • 721 • 2
Running on Zero MCP Featured 95 Breeze TTS 2 🎙 95 Bilingual TTS with voice design, cloning, and direction
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation Paper • 2502.03930 • Published Feb 6, 2025 • 3