Comparison
CosyVoice vs EmotiVoice
Verdict
Pick CosyVoice if cosyVoice is a Python-based multi-lingual large voice generation model. It supports extensive capabilities including fine-tuning, TTS (Text-To-Speech), and natural language generation; pick EmotiVoice if emotiVoice is an emotion-driven text-to-speech engine that emphasizes voice diversity and prompt control, allowing users to generate speech with varied emotions and speaking styles.
Markdown twin · CosyVoice alternatives · EmotiVoice alternatives
GraphCanon updated 2d
Trust & integrity
| Signal | CosyVoice | EmotiVoice |
|---|---|---|
| Maintenance | Steady (89d since push) As of 2d · github_public_v1 | Dormant (714d since push) As of 3w · github_public_v1 |
| Provenance | Not a fork · Organization account As of 2d · github_public_v1 | Not a fork · Personal account As of 3w · github_public_v1 |
| OSV dependency advisories | No lockfile (source not queried) As of 1mo · osv@v1 | No published findings from this source as of 2026-07-11 As of 1mo · osv@v1 |
| deps.dev advisories | Not queried deps.dev@v1 | Not queried deps.dev@v1 |
| OpenSSF Scorecard | Not queried openssf-scorecard@v1 | Not queried openssf-scorecard@v1 |
Tagline
- CosyVoice
- Multi-lingual large voice generation model with full-stack abilities for inference, training and deployment.
- EmotiVoice
- A Multi-Voice and Prompt-Controlled TTS Engine
Stars
- CosyVoice
- 23k
- EmotiVoice
- 8.5k
Forks
- CosyVoice
- 2.6k
- EmotiVoice
- 754
Open issues
- CosyVoice
- 719
- EmotiVoice
- 137
Language
- CosyVoice
- Python
- EmotiVoice
- Python
Adopt for
- CosyVoice
- CosyVoice is a Python-based multi-lingual large voice generation model. It supports extensive capabilities including fine-tuning, TTS (Text-To-Speech), and natural language generation.
- EmotiVoice
- EmotiVoice is an emotion-driven text-to-speech engine that emphasizes voice diversity and prompt control, allowing users to generate speech with varied emotions and speaking styles.
Persona
- CosyVoice
- -
- EmotiVoice
- -
Runtime
- CosyVoice
- -
- EmotiVoice
- -
License
- CosyVoice
- Apache-2.0
- EmotiVoice
- Apache-2.0
Last pushed
- CosyVoice
- May 25, 2026
- EmotiVoice
- Aug 13, 2024
Categories
- CosyVoice
- Inference & Serving, Model Training, Speech & Audio
- EmotiVoice
- Speech & Audio
Trust and health
Maintenance
- CosyVoice
- Steady (60%)
- EmotiVoice
- Dormant (18%)
Days since push
- CosyVoice
- 89d
- EmotiVoice
- 714d
Open issues (now)
- CosyVoice
- 719
- EmotiVoice
- 137
Stars delta
- CosyVoice
- +497 (30d)
- EmotiVoice
- Unknown
Open issues delta
- CosyVoice
- -37 (30d)
- EmotiVoice
- Unknown
Owner type
- CosyVoice
- Organization
- EmotiVoice
- User
OSV dependency advisories
- CosyVoice
- No lockfile (source not queried)
- EmotiVoice
- No published findings from this source as of 2026-07-11
Full report
- CosyVoice
- Trust report
- EmotiVoice
- Trust report
Shared compatibility
- Python · CosyVoice: Python runtime · EmotiVoice: Python runtime
Choose CosyVoice if…
- Tags unique to CosyVoice: audio-generation, cantonese, chatbot, chatgpt.
- Also covers Inference & Serving, Model Training.
- When you need support for multiple languages like Cantonese, Chinese, English, Japanese, and Korean.
When NOT to use CosyVoice
- If your project specifically requires fine-tuned performance in languages not supported by CosyVoice such as Arabic or Spanish.
- When strict real-time speech synthesis requirements are essential, as CosyVoice may face delays depending on the environment's computational power and model complexity.
Choose EmotiVoice if…
- Tags unique to EmotiVoice: ai, deep-learning, emotion, multispeaker.
- EmotiVoice ships Docker support for self-hosted deployment.
- When you need a TTS solution capable of handling multiple speaker voices in a single project
When NOT to use EmotiVoice
- For setups without access to an NVidia GPU, as EmotiVoice requires one for running its Docker image
- When deploying applications that have strict latency requirements and cannot accommodate the startup time of a Docker container or the processing demands of GPU computations
Explore
Sources
Every stat on this page traces to a dated GitHub sync, license file, enrichment field, or trust scan.
- GitHub stars (FunAudioLLM/CosyVoice) · observed Aug 23, 2026
- GitHub forks (FunAudioLLM/CosyVoice) · observed Aug 23, 2026
- Last push (FunAudioLLM/CosyVoice) · observed May 25, 2026
- License file (Apache-2.0) · observed Aug 23, 2026
- Decision facts (enrichment) · observed Jul 11, 2026
- Trust scan (lockfile / OSV) · observed Jul 11, 2026
- GitHub stars (netease-youdao/EmotiVoice) · observed Jul 29, 2026
- GitHub forks (netease-youdao/EmotiVoice) · observed Jul 29, 2026
- Last push (netease-youdao/EmotiVoice) · observed Aug 13, 2024
- License file (Apache-2.0) · observed Jul 29, 2026
- Decision facts (enrichment) · observed Jul 17, 2026
- Trust scan (lockfile / OSV) · observed Jul 11, 2026
GitHub stars on cards: CosyVoice 23k · EmotiVoice 8.5k (synced Aug 23, 2026).
Common questions
- What is the difference between CosyVoice and EmotiVoice?
- CosyVoice: Multi-lingual large voice generation model with full-stack abilities for inference, training and deployment.. EmotiVoice: A Multi-Voice and Prompt-Controlled TTS Engine. See the comparison table for live GitHub stats and shared categories.
- When should I choose CosyVoice over EmotiVoice?
- Choose CosyVoice over EmotiVoice when Tags unique to CosyVoice: audio-generation, cantonese, chatbot, chatgpt; Also covers Inference & Serving, Model Training; When you need support for multiple languages like Cantonese, Chinese, English, Japanese, and Korean.
- When should I choose EmotiVoice over CosyVoice?
- Choose EmotiVoice over CosyVoice when Tags unique to EmotiVoice: ai, deep-learning, emotion, multispeaker; EmotiVoice ships Docker support for self-hosted deployment; When you need a TTS solution capable of handling multiple speaker voices in a single project.
- When should I avoid CosyVoice?
- If your project specifically requires fine-tuned performance in languages not supported by CosyVoice such as Arabic or Spanish. When strict real-time speech synthesis requirements are essential, as CosyVoice may face delays depending on the environment's computational power and model complexity.
- When should I avoid EmotiVoice?
- For setups without access to an NVidia GPU, as EmotiVoice requires one for running its Docker image When deploying applications that have strict latency requirements and cannot accommodate the startup time of a Docker container or the processing demands of GPU computations
- Is CosyVoice or EmotiVoice more popular on GitHub?
- CosyVoice has more GitHub stars (22,870 vs 8,501). Stars measure visibility, not whether either tool fits your constraints.
- Are CosyVoice and EmotiVoice open source?
- Yes - both are open-source projects on GitHub (CosyVoice: Apache-2.0, EmotiVoice: Apache-2.0).
- Where can I find alternatives to CosyVoice or EmotiVoice?
- GraphCanon lists graph-backed alternatives at CosyVoice alternatives and EmotiVoice alternatives (CosyVoice markdown twin, EmotiVoice markdown twin), ranked by typed relationship edges rather than popularity votes.
- Is there a machine-readable version of this comparison?
- Yes. The markdown twin at this comparison mirrors this page for agents and LLM crawlers, with the same stats table and FAQ answers.
- Which is better maintained, CosyVoice or EmotiVoice?
- CosyVoice: Steady. EmotiVoice: Dormant. Compare maintenance labels, days since push, and release cadence in the trust section below - stars alone do not measure maintenance.
- Where are the full trust reports for CosyVoice and EmotiVoice?
- GraphCanon publishes per-repo trust reports with dated maintenance, provenance, and scan summaries: CosyVoice trust report; EmotiVoice trust report.