IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
🎙
217
Generate speech from text using a reference audio
VLMEvalKit Eval Results in video understanding benchmark
Track, rank and evaluate open LLMs and chatbots
Browse speech‑recognition model performance across languages