ChatTTS
✓ verified, sync 2026-08-20ChatTTS is a voice generation model designed for conversational scenarios. It is ideal for applications such as dialogue tasks for large language model assistants, as well as conve….
What it does.
ChatTTS is a voice generation model designed for conversational scenarios. It is ideal for applications such as dialogue tasks for large language model assistants, as well as conversational audio and video introductions. The model supports both Chinese and English, demonstrating high quality and naturalness in speech synthesis. This level of performance is achieved through training on approximately 100,000 hours of Chinese and English data. The project team plans to open-source a basic model trained with 40,000 hours of data, which will aid the academic and developer communities in further research and development.
Features
- Multi-language support (English and Chinese)
- High-quality and natural-sounding voice synthesis
- Dialog task compatibility for LLM assistants
- Open-source plan for a trained base model