Winnougan/INT4-Convrot-Comfy-Models
Updated • 58
Zero GPU Text-to-Speech using Fish Audio S2 Pro
Generate natural speech from text with Qwen3
Music Generation Foundation Model v1.5
Open-source autoregressive model with binary visual tokens.
Generate singing voice from lyrics and convert vocals
FireRed-Image-Edit-1.0
FireRed-Image-Edit × Qwen-Image-Edit-Rapid (Transformers)
MegaTTS 3 but with voice cloning!
Generate custom captions, tags, or prompts for any image
Generate a text prompt from an image
Generate spoken audio from text using selectable voices
A Step Towards Music Generation Foundation Model
Expressive Zeroshot TTS