netease-youdao/emotivoice

EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine

View on GitHub ↗Jump to charts ↓

Summary Information

Updated 35 minutes ago
Added to GitGenius on September 7th, 2026
Created on November 8th, 2023
Open Issues & Pull Requests: 137 (+0)
GitHub issues: Enabled
Number of forks: 755
Total Stargazers: 8,522 (+0)
Total Subscribers: 72 (+0)

Repository Insights (GitGenius)

Most active contributors

Sign in to see contributor activity.

Related repositories by overlapping contributors

No overlapping-contributor repos identified yet.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Issue API getrepoissuespagesummary failed: 429 Rate limit exceeded. Please try again later.

Detailed Description

EmotiVoice is a text-to-speech engine that synthesizes speech with emotional expression and multi-voice support.

The tool addresses the limitation of conventional TTS systems that produce neutral, emotionless speech. EmotiVoice solves this by enabling prompt-controlled synthesis, allowing users to specify emotions such as happiness, excitement, sadness, and anger when generating audio. The engine supports both English and Chinese, with access to a large voice library. It provides multiple interfaces for interaction: a web-based UI for interactive use and a scripting interface for batch processing.

The project suits developers and content creators who need expressive speech synthesis rather than flat, robotic output. It is particularly valuable for applications like audiobook production, interactive dialogue systems, and media where emotional tone matters. The tool offers an HTTP API with included free usage allowance, making it accessible for experimentation. A voice cloning capability allows users to create custom voices from personal audio data using provided recipes. The project does not position itself against specific alternatives in its README.

Development activity shows consistent feature expansion with recent additions including voice speed control in the OpenAI-compatible API, a native macOS application, and voice cloning functionality. The project actively incorporates community contributions and maintains responsiveness to user requests, as evidenced by merged pull requests addressing specific feature requests. The maintainers are tracking community feedback and explicitly welcoming input to guide future development priorities, including planned support for additional languages.