open-mmlab/amphion

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and...

View on GitHub ↗Jump to charts ↓Open shareable report

Summary Information

Updated 56 minutes ago
Added to GitGenius on September 6th, 2026
Created on November 15th, 2023
Open Issues & Pull Requests: 174 (+0)
GitHub issues: Enabled
Number of forks: 849
Total Stargazers: 10,282 (+0)
Total Subscribers: 92 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 35.8 hours
Mean response time: 15.2 days
90th percentile: 37.0 days
Tracked items: 122

How this project is maintained

Around half of the issues opened in the past year never receive a reply. 100% of open issues come from outside the core team, so the backlog reflects real-world use rather than internal planning. Only 5% of issues opened in the past year have been closed.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 129
New in 7 days: 0
Closed in 7 days: 0
Avg open age: 544 days
Stale 30+ days: 129
Stale 90+ days: 128

Recent activity

Opened in 7 days: 0
Closed in 7 days: 0
Comments in 7 days: 0
Events in 7 days: 0

Top labels

  • bug (38)
  • enhancement (24)
  • documentation (4)
  • Status: in progress (1)
  • macOS (1)
  • planned feature (1)

Most active issues this week

No issue events were indexed in the last 7 days.

Detailed Description

Amphion is a toolkit for audio, music, and speech generation that supports reproducible research and helps junior researchers and engineers enter the field of audio generation.

The toolkit addresses the challenge of implementing and understanding audio generation models by providing a unified platform for multiple generation tasks including text-to-speech, singing voice synthesis, voice conversion, accent conversion, singing voice conversion, and text-to-audio. A distinctive feature is its inclusion of model architecture visualizations designed to help junior researchers understand how classic models work. The toolkit also provides vocoders for producing high-quality audio signals and evaluation metrics to ensure consistent measurement across generation tasks.

Amphion suits researchers and engineers working on audio generation who need both implementation frameworks and educational resources. The toolkit is particularly valuable for those new to the field who benefit from architectural visualizations alongside working code. It supports individual generation tasks at different maturity levels, with text-to-speech, singing voice synthesis, voice conversion, accent conversion, singing voice conversion, and text-to-audio marked as supported, while text-to-music is noted as in development. The project includes large-scale dataset building capabilities for real-world applications like speech synthesis.

The project maintains active engagement across multiple platforms including model repositories and community channels. Development activity shows ongoing work across diverse generation tasks with varying levels of completion. The toolkit receives contributions addressing both core generation functionality and supporting infrastructure like vocoders and evaluation systems.