Open Source
Explore the latest AI open-source projects from GitHub and HuggingFace.
Explore the latest AI open-source projects from GitHub and HuggingFace.
Amphion is an open-source toolkit by OpenMMLab for audio, music, and speech generation research and production. It covers a broad range of generation tasks including text-to-speech, voice conversion, singing voice synthesis, text-to-audio, and vocoder training, with support for architectures like VALL-E, NaturalSpeech2, MaskGCT, and Vevo for zero-shot capabilities. Pre-trained models are available on HuggingFace and ModelScope, making it accessible for both researchers and engineers.
RVC-Project
The de facto open-source voice-conversion framework — trains a usable voice clone from 10 minutes of audio via top-1 feature retrieval, with a ~170ms real-time GUI (MIT).
myshell-ai
Instant voice cloning framework by MIT and MyShell with 36k+ GitHub stars, enabling zero-shot cross-lingual voice replication from just seconds of reference audio.
Jamie Pine
A local-first, open-source AI voice studio that clones voices, generates speech across 23 languages and 7 TTS engines, dictates into any app, and gives MCP agents a voice you own.
Anjok07
Deep-learning desktop GUI for vocal and stem separation from any audio track.