Local Text-to-Speech and Audiobook Makers

Have text read aloud in a natural voice, or turn an ebook into an audiobook, with local engines like Chatterbox and F5-TTS or with ebook2audiobook.

52 tools
A self-hosted text-to-speech API for Kokoro-82M. Generate speech locally on CPU, NVIDIA GPU or Apple Silicon, with multi-speaker audio and captions.

5.5KUpdated 3 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#Home Assistant integration#Multilingual#OpenAI-compatible API

An open-source video translation and AI dubbing tool for Windows, macOS and Linux, with local offline models or cloud APIs. Licensed under GPL-3.0.

19.2KUpdated 2 days agoGPL-3.0

macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual

Favicon of Kokoro

Kokoro

6 videos
Open-source text-to-speech model and library for local speech generation, with multilingual voices, Apache 2.0 licensing and Apple Silicon GPU support.

9.1KUpdated 1 year agoApache-2.0

macOS · Windows#Batch processing#Multilingual#ONNX

A mobile AI assistant that runs GGUF models on iOS and Android. Core chat works offline after a model download and needs no account.

8.5KUpdated 2 days agoMIT

iOS · Android#GGUF#Hugging Face integration#llama.cpp backend

A local LLM frontend that connects to local backends and cloud APIs, with lorebooks, image generation and voice. Open source under AGPL-3.0.

33.9KUpdated 2 weeks agoAGPL-3.0

macOS · Windows · Linux · Android · Docker · Web#Multilingual

Local audiobook converter turns ebooks into narrated audio with chapters and voice cloning. Runs on Windows, macOS and Linux under Apache 2.0.

20.3KUpdated 4 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning

Favicon of Lemonade

Lemonade

1 video
An open source local AI server for chat, image generation, and speech on Windows, macOS, and Linux, with APIs for apps and agents.

5.8KUpdated 2 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Docker#GGUF#Hugging Face integration#llama.cpp backend

Local AI audiobook converter that turns EPUB books into M4B audio with Kokoro voices. Runs on Windows, macOS and Linux under the MIT license.

8.7KUpdated 7 months agoMIT

macOS · Windows · Linux#Multilingual

Self-hosted speech API for transcription, translation and speech generation. Runs via Docker on CPU or GPU with faster-whisper, Kokoro and Piper.

3.7KUpdated 5 months agoMIT

Docker#OpenAI-compatible API#Streaming inference

Self-hosted AI model serving platform for Linux, Windows and macOS. Run language, speech and image models through an OpenAI-compatible API under Apache 2.0.

9.6KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input

Favicon of Applio

Applio

2 videos
Open-source AI voice conversion software for Windows, macOS and Linux. Convert audio, change your voice live and train models locally.

3.8KUpdated 2 days agoMIT

macOS · Windows · Linux · Web#Batch processing#Voice conversion

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

An offline speech-to-text and text-to-speech app for Linux and Sailfish OS, with local translation, voice typing and MPL-2.0 open-source licensing.

1.7KUpdated 1 week agoMPL-2.0

Linux#Multilingual#Works offline

Open-source text-to-speech toolkit for local speech generation, voice cloning and model training on Linux, macOS and Windows, licensed under MPL-2.0.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

Open-source AI podcast generator in Python with local HuggingFace models for transcripts and cloud speech services for multilingual audio.

6.6KUpdated 5 months agoApache-2.0

Docker · Web#Hugging Face integration#Multilingual#Multimodal input

An open-source subtitle editor for Windows, macOS and Linux, with Whisper transcription and translation through Ollama or LM Studio.

14.4KUpdated 1 day agoMIT

macOS · Windows · Linux#Batch processing#LM Studio integration#Multilingual

More in Voice, Speech and Music