Chatterbox-TTS-Server
Self-hosted server for generating expressive multilingual speech with voice cloning, audiobook-scale text processing, predefined voices, a Web UI, and.

Total stars
Last 7 days
1,406
HostingSelf-hostable
LicenseMIT
LanguagePython
Last commitMay 26, 2026
Latest releaseMay 11, 2026
Release frequency1.7 per year
Contributors21
Stars1,406
Forks343
Issues49
What it does
Chatterbox TTS Server runs Resemble AI’s Chatterbox model family behind a modern Web UI and OpenAI-compatible API. Users can enter text, choose among Original, Multilingual, and Turbo engines, adjust generation parameters, use predefined voices or reference audio for cloning, and generate expressive speech with tags such as [laugh] and [cough]. It supports 23 languages through Chatterbox Multilingual and can process long text by chunking and joining segments for audiobook generation.
The project is self-hostable and documents automated local installation, Windows portable mode, Google Colab use, and Docker deployment. It runs with NVIDIA CUDA, AMD ROCm, Apple Silicon MPS, or CPU fallback, and exposes API endpoints including streaming text-to-speech and voice listing. The documented scope centers on speech synthesis and voice cloning; transcription, dubbing, music, sound-effects, and the
Capabilities
#Text-to-speech voiceover production
#Voice cloning and design
#Audio and agent APIs
Built with
Tools similar to Chatterbox-TTS-Server
GPT-SoVITS
Audio Editing
ChatTTS
Audio Editing
Ebook2audiobook
Audio Editing
Voice-pro
Audio Editing
VoiceStudio
Audio Editing
EmotiVoice
Audio Editing