Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
vmdb
's Collections
NVFP4
SOTA Coding
Leaderboards
DGX Spark - GB10
Vision
Language-Translation
Text-to-image
Image-Editing
Text-to-video
Image-to-video
Video-editing
Text-to-speech
Speech-to-text
Embedding models
8 GB VRAM
12 GB VRAM
16 GB VRAM
24 GB VRAM
Content Safety
Text-to-speech
updated
19 days ago
Upvote
-
Sort: Collection
mistralai/Voxtral-4B-TTS-2603
Text-to-Speech
•
Updated
Mar 31
•
1.17k
•
904
ResembleAI/chatterbox
Text-to-Speech
•
Updated
Jun 10
•
1.91M
•
•
1.77k
ResembleAI/chatterbox-turbo
Text-to-Speech
•
Updated
Dec 15, 2025
•
•
681
ResembleAI/chatterbox-turbo-ONNX
Text-to-Speech
•
Updated
Dec 15, 2025
•
3.43k
•
84
hexgrad/Kokoro-82M
Text-to-Speech
•
Updated
Apr 10, 2025
•
12.3M
•
•
6.75k
onnx-community/Kokoro-82M-v1.0-ONNX
Text-to-Speech
•
Updated
Feb 8, 2025
•
1.31M
•
253
fishaudio/s2-pro
Text-to-Speech
•
5B
•
Updated
Mar 11
•
415k
•
1.28k
Zyphra/Zonos-v0.1-transformer
Text-to-Speech
•
2B
•
Updated
Jun 3, 2025
•
142k
•
•
435
stepfun-ai/Step-Audio-EditX
4B
•
Updated
Feb 14
•
14.4k
•
138
nvidia/NVIDIA-NemotronLabs-VoiceChat-11B
11B
•
Updated
7 days ago
•
2.72k
•
432
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections