🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
-
Updated
May 21, 2026 - Python
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
A ComfyUI integration for FireRedTTS‑2, a real-time multi-speaker TTS system enabling high-quality, emotionally expressive dialogue and monologue synthesis. Leveraging a streaming architecture and context-aware prosody modeling, it supports natural speaker turns and stable long-form generation, ideal for interactive chat and podcast applications.
FireRedTTS3 for ComfyUI: multilingual zero-shot voice cloning, voice design, speech editing, Whisper transcripts, AIMDO DynamicVRAM, with bf16 / INT8 compatibility
ComfyUI nodes for FireRedTTS3: zero-shot voice cloning in 24 languages and 21 Chinese dialects, instruction-based voice design, speech editing, and DynamicVRAM support.
To associate your repository with the fireredtts topic, visit your repo's landing page and select "manage topics."