GPT-SoVITS
GPT-SoVITS is a few-shot voice cloning and TTS WebUI that fine-tunes a speaker's voice from one minute of audio and supports zero-shot synthesis from a five-second sample.
What it solves
GPT-SoVITS is a few-shot voice cloning and TTS WebUI that fine-tunes a speaker's voice from one minute of audio and supports zero-shot synthesis from a five-second sample.
The project context on this page is free to read. The original GitHub repository remains the source of truth; sign in only when you want to save or join the discussion.
Community notes