Audiblez turns EPUB e-books into M4B audiobooks using Kokoro-82M text-to-speech on your own computer. It's for readers who want spoken versions of their books and control over the narrator, reading speed and sections included. The app is open source under the MIT license.
You can use a graphical interface or a command-line tool. Both serve the same purpose: converting an e-book into audio you can play in VLC or an audiobook player. It also produces separate WAV files for chapters, so the output includes individual audio files as well as the complete book.
Voice choices cover American and British English, Spanish, French, Hindi, Italian, Japanese, Brazilian Portuguese and Mandarin Chinese. You can choose a narrator and adjust the speech speed during conversion. Chapter selection lets you leave out sections you don't want read aloud, including through an interactive picker in the command-line tool.
Audiblez runs on Windows, macOS and Linux. Speech generation uses the CPU by default, with optional NVIDIA GPU acceleration through CUDA and PyTorch. An Apple Silicon Mac can run it on the CPU, but it doesn't support Apple Silicon GPU acceleration. Google Colab is another supported place to run conversion with CUDA, though that uses cloud hardware rather than your computer.
The audiobook conversion relies on Kokoro for speech, espeak-ng for speech processing and FFmpeg to assemble the final M4B file.
Claim this page and we'll verify you by hand. audiblez gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find audiblez?Promote it
Something wrong or outdated on this page?
1.5KUpdated 4 months agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API
Chatterbox TTS Server runs Resemble AI's speech models on your own computer or server, with a browser interface and an OpenAI-compatible API. It's for people producing narration and audiobooks, or developers adding speech to voice agents and other apps. The project is open source under the MIT license.
2.3KUpdated 4 months agoMPL-2.0
macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning
20.3KUpdated 4 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning
62.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection
5.5KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#Home Assistant integration#Multilingual#OpenAI-compatible API
11.2KUpdated 1 month ago
macOS · Windows · Linux · iOS · Android · Web#Multilingual#Streaming inference
Coqui TTS (idiap fork) is a local text-to-speech library for developers and speech researchers who want pretrained voices or tools to train their own models. It builds on coqui-ai/TTS, continuing the original unmaintained project. The Python toolkit is open source under the Mozilla Public License 2.0 (MPL-2.0).
ebook2audiobook turns non-DRM ebooks into narrated audio with chapters and metadata, for readers who want audio editions of their own books. It runs locally on Windows, macOS and Linux, with Docker support and a browser interface built with Gradio. It's open source under Apache 2.0.
GPT-SoVITS is a local text-to-speech and voice cloning tool. It can generate speech from a short reference recording or fine-tune a model for a custom voice. The source code uses the MIT license.
Kokoro-FastAPI runs the Kokoro-82M speech model on your own machine or server and exposes an OpenAI-compatible speech API. It's for developers adding local text-to-speech to assistants, reading apps or audiobook workflows. Speech generation runs locally, and the API doesn't require an OpenAI account.
Moonshine is an on-device AI toolkit for developers building voice agents and applications that listen and speak. It combines speech to text, intent recognition and text to speech in one library. Voice processing stays on the device, and you don't need an account or API keys.