Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Log In
Sign Up
he-shuwei
/
M2SE-VTTS
like
1
Text-to-Speech
English
visual-tts
speech-synthesis
diffusion
spatial-audio
arxiv:
2412.11409
License:
mit
Model card
Files
Files and versions
xet
Community
main
M2SE-VTTS
/
bigvgan
449 MB
Ctrl+K
Ctrl+K
1 contributor
History:
2 commits
he-shuwei
Upload bigvgan/g_00076000 with huggingface_hub
aff1cee
verified
19 days ago
config.json
Safe
1.37 kB
Upload bigvgan/config.json with huggingface_hub
19 days ago
g_00076000
449 MB
xet
Upload bigvgan/g_00076000 with huggingface_hub
19 days ago