Repo Voice & Audio Limited info

myshell-ai/OpenVoice

OpenVoice is an MIT-licensed instant voice cloning model from MIT and MyShell, with multilingual support and control over voice style.

  • 37.8k GitHub stars
  • Python
  • ⚖️ MIT
  • 🎯 Intermediate
myshell-ai/OpenVoice preview image

What it is

OpenVoice is an instant voice cloning system developed by researchers at MIT, Tsinghua University and MyShell. It clones a reference speaker's tone color and generates speech in multiple languages and accents. V2, released in April 2024, adds better audio quality and native multilingual support. It has powered the voice cloning feature on myshell.ai since May 2023.

Who it's for

  • Developers who need voice cloning in their applications
  • Researchers working on voice cloning and cross-lingual speech generation
  • Teams that need a commercially usable voice cloning model under the MIT License

Pros & cons

Pros

  • Pro:Accurate tone color cloning that reproduces the reference voice in multiple languages and accents
  • Pro:Granular style control over emotion, accent, rhythm, pauses and intonation
  • Pro:Zero-shot cross-lingual cloning: neither the generated language nor the reference language has to appear in the training dataset
  • Pro:MIT licensed for both V1 and V2, free for commercial and research use

Cons

  • Con:The README has no install or setup commands; it only points to docs/USAGE.md for instructions
  • Con:Native language support in V2 is limited to English, Spanish, French, Chinese, Japanese and Korean

Images