ChatTTS
ChatTTS generates expressive spoken audio from text for dialogue applications, with multilingual support, multiple speakers, streaming, and fine-grained.

Total stars
Last 7 days
39,767
HostingSelf-hostable
LicenseAGPL-3.0
LanguagePython
Last commitApr 10, 2026
Latest releaseApr 10, 2026
Release frequency3.2 per year
Contributors59
Stars39,767
Forks4,258
Issues61
What it does
ChatTTS is an open-source generative speech model designed for dialogue scenarios such as LLM assistants. It synthesizes expressive speech in English and Chinese, supports multiple speakers, and provides control over prosodic elements including laughter, pauses, and interjections. Streaming audio generation and sampled speaker embeddings are documented capabilities.
The repository includes a WebUI, command-line inference, Python usage examples, and audio export through torchaudio. It can be installed from PyPI, GitHub, or a local checkout and run with local Python dependencies and suitable hardware; the README cites at least 4GB of GPU memory for a 30-second clip. The code is AGPLv3+, while the released model is CC BY-NC 4.0 for academic and research use.
Capabilities
#Text-to-speech voiceover production
Built with
Tools similar to ChatTTS
GPT-SoVITS
Audio Editing
Ebook2audiobook
Audio Editing
Voice-pro
Audio Editing
VoiceStudio
Audio Editing
EmotiVoice
Audio Editing
Abogen
Audio Editing