
Immersive Translate is an AI translation extension and mobile app that keeps original text alongside its translation. It's aimed at students, researchers and people who read foreign-language material for work. The bilingual page layout lets readers compare passages without replacing the source text.
It runs in Chrome, Edge, Firefox and Safari on macOS, with apps for iOS and Android. It connects to translation services including ChatGPT, DeepL, DeepSeek and Gemini. Its official Ollama guide also documents a local model backend. Translation processing follows the selected service; cloud engines receive the content sent for translation, while Ollama can run the selected model on your own hardware.
Document translation preserves PDF formatting and can produce bilingual or translated-only output. OCR handles scanned PDFs. It also accepts EPUB books, DOCX documents, HTML, plain text, Markdown and subtitle files, covering both reading material and work documents.
For video, it displays bilingual subtitles on YouTube, Netflix and Prime Video, and can translate videos that lack original subtitles. Meeting translation works with Zoom, Google Meet and Microsoft Teams. Image and comic translation covers web images and images supplied from your device, using OCR and inpainting to retain their layout and visual style.
Context-aware translation and customizable terminology libraries support specialized language across pages, PDFs and video subtitles. Readers can translate selected text or individual paragraphs within a page, while input-box translation helps with writing messages and searches in another language.
Claim this page with an email at immersivetranslate.com. Immersive Translate gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Immersive Translate?Promote it
Something wrong or outdated on this page?
14.4KUpdated 24 hours agoMIT
macOS · Windows · Linux#Batch processing#LM Studio integration#Multilingual
Subtitle Edit is an MIT-licensed subtitle editor for Windows, macOS and Linux. It's for people creating captions, translating dialogue or fixing subtitles that don't match the video. Its core editing, conversion and video playback work offline on your device, with optional AI tools for transcription and translation.
sindresorhus.comOn-Device and In-Browser AI
macOS · iOS#Batch processing#Multilingual
goodsnooze.gumroad.comDictation and Voice Typing
macOS · iOS#Batch processing#Multilingual#Ollama integration
21.8KUpdated 1 week agoMIT
macOS · Windows · Linux#Hugging Face integration#Multilingual#Speaker diarization
20.3KUpdated 4 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning
13.1KUpdated 2 years agoApache-2.0
macOS · Windows#Batch processing#Hugging Face integration#Multilingual
Aiko is a paid, native transcription app for macOS, iOS and visionOS that processes speech on your device with OpenAI's Whisper model. It's for people turning meetings, lectures or other recordings into text while keeping the audio local, including sensitive recordings.
MacWhisper is a native macOS transcription app for people working with interviews, lectures, meetings and other recorded audio. It runs speech recognition on your own Mac, so local transcription keeps audio on your device. It also offers cloud transcription through services such as OpenAI, ElevenLabs and Deepgram, which send audio off your machine.
Buzz transcribes and translates speech on your own computer using OpenAI's Whisper. It's for people who need transcripts or subtitles from recordings, plus live captions from a microphone. Local transcription works offline; the optional OpenAI Whisper API sends audio to a cloud service.
ebook2audiobook turns non-DRM ebooks into narrated audio with chapters and metadata, for readers who want audio editions of their own books. It runs locally on Windows, macOS and Linux, with Docker support and a browser interface built with Gradio. It's open source under Apache 2.0.
insanely-fast-whisper is a command-line tool for people who want to transcribe audio on their own hardware, with a focus on processing long recordings quickly. It runs OpenAI's Whisper locally on NVIDIA GPUs or Apple Silicon Macs, including support for Windows with CUDA. The project is open source under the Apache 2.0 license.