Winnougan/INT4-Convrot-Comfy-Models
Updated β’ 63
Generate natural speech from your text instantly
Music Generation Foundation Model v1.5
Open-source autoregressive model with binary visual tokens.
Generate singing voice from lyrics and convert vocals
FireRed-Image-Edit-1.0
FireRed-Image-Edit Γ Qwen-Image-Edit-Rapid (Transformers)
Generate high-quality images from text prompts
MegaTTS 3 but with voice cloning!
Generate custom captions, tags, or prompts for any image
Generate a text prompt from an image
Generate spoken audio from text using selectable voices
A Step Towards Music Generation Foundation Model
Expressive Zeroshot TTS