SoniTranslate is a local AI video dubbing app for creators and translators who need speech in another language to follow the timing of the original video. Its Gradio browser interface brings transcription, translation and speech generation together, with speaker detection for recordings that contain multiple voices. Local installation is tested on Linux, and it can use an NVIDIA GPU or run in CPU mode.
Whisper handles transcription, while speech options include Piper, Coqui XTTS, BARK and Facebook MMS. XTTS can clone a voice from a short recording; OpenVoice and RVC provide other voice imitation options. You can edit translated subtitles and speaker assignments, adjust speech speed and volume, and separate vocals from other audio. Supported languages include English, Japanese, Arabic, Ukrainian and Simplified or Traditional Chinese, though some languages support translation without transcription.
Outputs include dubbed video, separate audio, subtitles by speaker, and video with subtitles alone. It supports SRT and ASS subtitles, including subtitles burned into the video. Batch processing covers multiple files and full YouTube playlists. Document translation also extends to PDF videobooks that display images from the PDF.
The app runs locally, but optional OpenAI transcription, translation and speech generation send content to a cloud API. Colab and an online demo offer hosted alternatives. Pyannote speaker detection requires a Hugging Face account and acceptance of its model terms. The code is open source under Apache 2.0; individual models and weights may carry commercial restrictions.
Claim this page and we'll verify you by hand. SoniTranslate gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find SoniTranslate?Promote it
Something wrong or outdated on this page?
19.2KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual
pyVideoTrans translates spoken audio into another language and produces a video with translated subtitles and AI dubbing. It's for people adapting videos for audiences in other languages who want control over which parts run locally. It recognizes speech directly, so the original video doesn't need subtitles.
5.4KUpdated 1 day agoMIT
macOS · Windows · Linux#Batch processing#MCP#Multilingual
13KUpdated 3 months agoGPL-3.0
macOS · Windows · Linux · Web#Hugging Face integration#Multilingual#Quantization
5.6KUpdated 2 weeks agoApache-2.0
macOS · Windows · Linux · Web#Multilingual#OpenAI-compatible API#Resumable workflows
18.6KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#MLX#Multilingual
549Updated 2 weeks agoMIT
Windows · Linux#Multilingual#Multimodal input#Streaming inference
SmartSub is a free, open-source desktop app for people who subtitle recordings or adapt videos into other languages. It combines local transcription, translation, subtitle editing and AI dubbing on Windows, macOS and Linux. Each stage also works independently.
Voice-Pro brings transcription, voice cloning and multilingual dubbing into a locally run Gradio web app. It's for podcasters, video creators and developers who want to process recordings and generate speech in one interface. The software is free and open source under GPL-3.0.
YouDub-webui is a self-hosted video translation and dubbing app for creators and small teams who want to process media on their own hardware. It accepts YouTube and Bilibili links or local video files, then produces translated subtitles, cloned-voice dubbing, or both. English-to-Chinese dubbing for YouTube is its most established workflow; it also supports Chinese-to-English dubbing for Bilibili.
VideoLingo is a self-hosted video translation app for creators and educators who need bilingual subtitles or dubbed versions of their videos. It brings transcription, translation and subtitle timing into one browser interface, with dubbing as an optional output. The project is open source under Apache 2.0; a separate hosted service offers subtitle translation and dubbing.
Whispering Tiger turns audio on your computer into live transcripts and translations for VRChat and streaming overlays. It's free, open source software under the MIT license, aimed at people who want captions or translated conversations while gaming or broadcasting. Processing stays on your machine, and it works offline once you've downloaded the models.