Amphion

AI Audio & Voicefreemium★ 10,303 GitHub starsGitHub

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.

Visit Amphion

Alternatives to Amphion

Transformers

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference a

freemium★ 166.4k

Deer Flow

An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message ga

free★ 82.7k

GPT SoVITS

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

freemium★ 61.9k

TTS

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

freemium★ 46.0k

ChatTTS

A generative speech model for daily dialogue.

freemium★ 39.9k

VoxCPM

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

freemium★ 37.8k