RVC-Boss/GPT-SoVITS
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
Appears on
Quick read
Latest capture 2026-08-03 03:09
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +288
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Large-scale LLM inference engine
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
We turn natural language descriptions of behaviors into machine-executable code
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
A family of lightweight multimodal models.