Alternatives to OpenVoice
Open source instant voice cloning model with flexible style control. The listings below can replace it for an important use case. Each note says what changes if you switch.
The original
OpenVoice
Open source instant voice cloning model with flexible style control.
Replacements
Listings that take over the same core job as OpenVoice.
Chatterbox
Open source text-to-speech model with zero-shot voice cloning.
Chatterbox is an MIT-licensed text-to-speech model with zero-shot voice cloning from a very short sample, but it has no dedicated GUI and is mainly used through code.
Coqui TTS
Deep learning toolkit for training and running text-to-speech models.
Coqui TTS is an MPL-2.0 toolkit with pretrained models in many languages and support for full training or fine-tuning, though it has not been updated since mid-2024.
Bark
Open source generative audio model for expressive multilingual speech.
Bark is an MIT generative audio model that produces expressive multilingual speech with nonverbal sounds like laughter, but it is slower and more resource-intensive than lightweight TTS options.
Tortoise TTS
Open source multi-voice text-to-speech system tuned for quality.
Tortoise TTS is an Apache-2.0 multi-voice system that emphasises realistic prosody and intonation, at the cost of notably slower generation and no updates since late 2024.
Kokoro
Small, fast open-weight text-to-speech model you can run locally.
Kokoro is a small Apache-licensed text-to-speech model that is fast and cheap to run locally, though it is used as a Python library without a graphical interface.
F5-TTS
Open-source text-to-speech model that can clone voices, with a local Gradio interface for running it.
F5-TTS clones voices from a short reference recording and adds a local Gradio interface, as a large active project that needs Python and ideally a capable GPU.
GPT-SoVITS
A voice cloning and text-to-speech toolkit that can train a voice from about one minute of audio.
GPT-SoVITS trains a custom voice from about a minute of audio with a local web UI on Windows and Linux, but it is not open source.
Applio
A free, open-source AI voice conversion suite for covers, custom voice training and real-time voice changing.
Applio is an MIT licensed voice conversion suite with a graphical app for covers, real-time voice changing and model training, rather than instant cloning from code.
Also worth comparing
These listings name OpenVoice as their own alternative, so the relationship runs both ways.
Piper
A fast neural text-to-speech engine that runs locally, even on modest hardware like a Raspberry Pi.
OpenVoice adds instant voice cloning with control over emotion and accent under the MIT license, though it is mainly used through code and has not been updated since April 2025.
Similar software
Related functionality, not a direct replacement.