Local AI Voice Cloning and Voice Changers

Clone a voice from a recording or change yours in real time with GPT-SoVITS or RVC WebUI. Use them only with the speaker's consent.

30 tools
Favicon of Chatterbox

Chatterbox

3 videos
An open-source text-to-speech model family that runs on your own hardware, clones voices from short clips, and supports offline deployment.

26.6KUpdated 2 months agoMIT

Linux#Multilingual#Voice cloning#Voice conversion

Favicon of IndexTTS

IndexTTS

1 video
Local text-to-speech software clones voices from one audio clip, supports five languages, and provides separate controls for emotion and speaking speed.

24.2KUpdated 1 day ago

Windows · Linux · Web#Hugging Face integration#Multilingual#Multimodal input

Self-hosted text-to-speech with voice cloning, multilingual speech and emotion control. Code and weights use the FISH AUDIO RESEARCH LICENSE.

32.9KUpdated 2 weeks ago

#Batch processing#Multilingual#Multimodal input

Local AI voice conversion software for Windows, Linux and Apple Silicon Macs. Converts speech and singing from a short voice sample; GPL-3.0 and archived.

3.9KUpdated 1 year agoGPL-3.0

macOS · Windows · Linux · Web#Hugging Face integration#Streaming inference#Voice conversion

An open-source singing voice conversion framework that runs fully offline with user-trained models. Licensed under AGPL-3.0; archived and no longer maintained.

28.1KUpdated 3 years agoAGPL-3.0

#Hugging Face integration#ONNX#Voice conversion

A self-hosted text-to-speech server using Piper and Coqui XTTS v2, with voice cloning and an OpenAI-compatible API. Archived and no longer maintained.

857Updated 2 years agoAGPL-3.0

macOS · Windows · Linux · Docker#Multilingual#ONNX#OpenAI-compatible API

Higgs Audio V2, now Higgs TTS 2, is a downloadable speech model for expressive narration, multilingual dialogue and voice cloning.

8.4KUpdated 4 months agoApache-2.0

#Batch processing#Hugging Face integration#Multilingual

Open-source text-to-speech software that runs locally, generates English speech and supports voice cloning. MIT licensed, with downloadable models.

4.7KUpdated 1 year agoMIT

#Hugging Face integration#Multilingual#Voice cloning

Open-source text-to-speech model with voice cloning. Runs locally on Linux and macOS under Apache 2.0, with a hosted audio playground also available.

7.2KUpdated 2 years agoApache-2.0

macOS · Linux · Docker · Web#Multilingual#Voice cloning

An open-source text-to-speech model you can run locally, with MIT-licensed Python code, pretrained English voices and adaptation to unfamiliar speakers.

6.4KUpdated 3 years agoMIT

Windows#Hugging Face integration#Multilingual#Voice cloning

Favicon of XTTS v2

XTTS v2

1 video
Local text-to-speech model with voice cloning and streaming audio, available through Coqui TTS on Linux, macOS and Windows.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

Open-source text-to-speech software for local voice cloning and streaming speech generation, with Apache 2.0 licensing and NVIDIA GPU deployment.

23.8KUpdated 4 months agoApache-2.0

Linux · Docker · Web#Hugging Face integration#Multilingual#Streaming inference

Open-source local text-to-speech built on Qwen2.5, with Chinese and English voice cloning, adjustable voices, and an Apache 2.0 license.

11KUpdated 1 year agoApache-2.0

macOS · Windows · Linux · Web#Hugging Face integration#Multilingual#Voice cloning

Open-source text-to-speech software generates custom voices locally on NVIDIA GPUs or Apple Silicon, with Docker support and an Apache 2.0 license.

14.9KUpdated 2 years agoApache-2.0

macOS · Windows · Docker#Streaming inference#Voice cloning

Open-source text-to-speech built on Llama, with local inference, voice cloning and streaming audio. Uses Apache 2.0; Baseten offers cloud hosting.

6.3KUpdated 10 months agoApache-2.0

#Hugging Face integration#llama.cpp backend#LoRA

Open-source text-to-speech model for local English dialogue generation, with voice cloning, NVIDIA GPU inference and an Apache 2.0 license.

19.4KUpdated 10 months agoApache-2.0

Docker · Web#Hugging Face integration#Multimodal input#Voice cloning

Local text-to-speech software generates speech from a reference voice, supports English and Chinese, and runs with NVIDIA GPUs. Code uses the MIT license.

15.3KUpdated 1 week agoMIT

Docker · Web#Multilingual#Voice cloning

Local AI voice conversion software for Windows and Linux. Train custom voices, convert recordings, and change your voice live with MIT-licensed code.

38.6KUpdated 2 months agoMIT

Windows · Linux · Web#Hugging Face integration#ONNX#Voice conversion

Local text-to-speech software built on Coqui TTS, with XTTSv2 models, voice fine-tuning and integrations for SillyTavern and Text-generation-webui.

2.4KUpdated 2 years agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Hugging Face integration

An open-source React Native library that runs GGUF models on iOS and Android through llama.cpp, with GPU acceleration and image and audio understanding.

1KUpdated 3 days agoMIT

iOS · Android#GGUF#llama.cpp backend#Multilingual

A self-hosted text-to-speech server that connects Piper to Home Assistant through Wyoming, with custom ONNX voices and optional NVIDIA GPU support.

214Updated 3 weeks agoMIT

Linux · Docker · Web#Home Assistant integration#Hugging Face integration#Multilingual

A self-hosted text-to-speech and audio generation interface with model extensions, Docker support and an OpenAI-compatible speech API. MIT licensed.

3.3KUpdated 3 weeks agoMIT

Windows · Docker · Web#OpenAI-compatible API

Local text-to-speech and voice cloning software with a browser interface, multilingual speech generation, and an MIT license. Runs on Windows, Linux and macOS.

62.2KUpdated 1 month agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection

Open-source voice cloning uses a reference recording to generate multilingual speech with style control. The Python project uses the MIT license.

37.7KUpdated 1 year agoMIT

#Multilingual#Voice cloning

An AI voice changer that converts speech live on Windows, Apple Silicon Macs and Linux, with Beatrice and RVC support and optional remote processing.

21.1KUpdated 4 days ago

macOS · Windows · Linux · Docker#ONNX#Voice conversion

An open-source video translation and AI dubbing tool for Windows, macOS and Linux, with local offline models or cloud APIs. Licensed under GPL-3.0.

19.2KUpdated 2 days agoGPL-3.0

macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual

Local audiobook converter turns ebooks into narrated audio with chapters and voice cloning. Runs on Windows, macOS and Linux under Apache 2.0.

20.3KUpdated 4 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning

Favicon of Applio

Applio

2 videos
Open-source AI voice conversion software for Windows, macOS and Linux. Convert audio, change your voice live and train models locally.

3.8KUpdated 2 days agoMIT

macOS · Windows · Linux · Web#Batch processing#Voice conversion

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

Open-source text-to-speech toolkit for local speech generation, voice cloning and model training on Linux, macOS and Windows, licensed under MPL-2.0.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

More in Voice, Speech and Music