EmotiVoice
EmotiVoice is an open-source text-to-speech app for expressive English and Chinese speech, prompt-controlled styles, voice cloning, batch generation, and APIs.
Total stars
Last 7 days
8,514
HostingSelf-hostable
LicenseApache-2.0
LanguagePython
Last commitAug 13, 2024
Latest releaseDec 28, 2023
Release frequency0.7 per year
Contributors13
Stars8,514
Forks757
Issues138
What it does
EmotiVoice is an open-source text-to-speech engine that generates English and Chinese speech with more than 2,000 voices. Its defining workflow uses prompts to control emotion and style, including happy, excited, sad, and angry delivery. The README also documents personal-data voice cloning, adjustable voice speed, an interactive web interface, batch scripting, and an OpenAI-compatible TTS API.
Users can run the provided Docker image with an NVIDIA GPU, install the project through Conda and pip, or use the Streamlit demo page locally. The project includes pretrained-model setup and command-line inference workflows, and an HTTP API is available for application integration. It is released under the Apache-2.0 license.
Capabilities
#Text-to-speech voiceover production
#Voice cloning and design
#Audio and agent APIs
Built with
Tools similar to EmotiVoice
GPT-SoVITS
Audio Editing
ChatTTS
Audio Editing
Ebook2audiobook
Audio Editing
Voice-pro
Audio Editing
VoiceStudio
Audio Editing
Abogen
Audio Editing