Open source audio-generation projects
Every project in the registry tagged audio-generation, ranked by real GitHub adoption.
Self-hosted AI runtime for text, voice, vision, and agents
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engin
A framework for efficient model inference with omni-modality models
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.
Genblaze is an open source Python SDK for orchestrating generative AI media pipelines across video, audio, and image providers with built in provenance for ever
Related tags
Frequently asked questions
How many open source audio-generation projects are there?
This registry tracks 6 projects tagged audio-generation, with 91,731 GitHub stars between them. The most-adopted is LocalAI at 49,144 stars.
Are these audio-generation projects free to use?
Yes — 6 of the 6 carry an explicit open-source licence across 2 distinct licences, so there is no licence fee. Where a project also sells a hosted or enterprise version, the self-hosted path remains free.
Which audio-generation project should I choose?
The list above is ranked by GitHub stars, but stars measure attention rather than fit. Check three things on each card: the licence (permissive versus copyleft), the language it is written in, and the last-push date — a high-star project that has not been pushed in a year is a liability.
Are these audio-generation projects still maintained?
4 of the 6 were pushed in the last 90 days, and every card shows its exact last-push date so you can see the rest. Sort your shortlist by that date before committing to a migration.