jickman125/firered-tts3
0
๐ฅ FireRedTTS3
Interactive demo of FireRedTeam/FireRedTTS3, a unified speech generation and editing model (Qwen3-1.7B backbone + flow-matching head over the continuous RedAE audio autoencoder, 24 kHz).
Three tabs:
Text normalization runs locally via wetext (Chinese / English); automatic language routing uses fastText lid.176. The optional LLM-based normalizer from the upstream repo is disabled here (it requires external API credentials).
Credits
- Model and inference code: FireRedTeam/FireRedTTS3, Apache-2.0.
examples/en_prompt.wav: from OpenBMB/VoxCPM (examples/example.wav), Apache-2.0.examples/zh_prompt.wav: from FireRedTeam/FireRedTTS2 (examples/chat_prompt/zh/S2.flac), Apache-2.0.
Voice cloning is provided for academic research purposes only โ do not use it for impersonation or any illegal activity.
