2 items
Clone a voice and train a text-to-speech model from about one minute of voice data using a WebUI.
Clone a voice from a reference clip and generate speech in multiple languages with control over emotion, accent, and rhythm.