TranscriptionSuite is a local speech-to-text app for people transcribing lectures, recording conversations or dictating into other apps. It's open source under GPL-3.0, with desktop apps for Windows, macOS and Linux. Transcription works offline after the initial app and model downloads, keeping audio and transcripts on your own hardware.
It records microphone or system audio and imports audio and video files. Long recordings get a rolling transcript preview; Live Mode provides sentence-by-sentence dictation through faster-whisper or whisper.cpp. System-wide shortcuts let you start and stop recording and paste text at the cursor. You can export plain text or SRT and ASS subtitles. Speaker labels work with supported models, but aren't available during live dictation.
Model choices include Whisper, NVIDIA NeMo Parakeet and Canary, SenseVoice, VibeVoice-ASR and whisper.cpp. Hardware support covers CPU processing, NVIDIA CUDA and AMD or Intel GPUs through Vulkan. Apple Silicon Macs run MLX models natively with Metal, without Docker; Windows, Linux and Intel Macs use a Docker or Podman server. Intel Macs use the CPU. Language and translation support depend on the model.
The Audio Notebook keeps recordings in a calendar view with full-text search. Its chat assistant connects to OpenAI-compatible providers, including local LM Studio and Ollama servers. Choosing a cloud provider sends note content to that service. PyAnnote speaker labeling requires a Hugging Face account token; SenseVoice, VibeVoice and Apple Silicon's Sortformer offer alternatives without one. You can also connect to a self-hosted transcription server over LAN or Tailscale, expose transcription to Open-WebUI through an OpenAI-compatible API, and send completed transcripts to automations through webhooks.
Claim this page and we'll verify you by hand. TranscriptionSuite gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find TranscriptionSuite?Promote it
Something wrong or outdated on this page?
goodsnooze.gumroad.comDictation and Voice Typing
macOS · iOS#Batch processing#Multilingual#Ollama integration
MacWhisper is a native macOS transcription app for people working with interviews, lectures, meetings and other recorded audio. It runs speech recognition on your own Mac, so local transcription keeps audio on your device. It also offers cloud transcription through services such as OpenAI, ElevenLabs and Deepgram, which send audio off your machine.
1.7KUpdated 2 weeks agoMPL-2.0
Linux#Multilingual#Works offline
21.8KUpdated 1 week agoMIT
macOS · Windows · Linux#Hugging Face integration#Multilingual#Speaker diarization
1.6KUpdated 3 weeks agoGPL-2.0
macOS · Windows · Linux#Hugging Face integration#Multilingual#Streaming inference
8.9KUpdated 22 hours agoMIT
macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual
19.2KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual
Speech Note combines offline dictation, reading aloud and translation in a desktop app for Linux and Sailfish OS. It's for people who want to take multilingual notes, type by voice or listen to text without sending their words to a cloud service. Speech and text processing stay on your device; models are downloaded separately through the app's graphical browser.
Buzz transcribes and translates speech on your own computer using OpenAI's Whisper. It's for people who need transcripts or subtitles from recordings, plus live captions from a microphone. Local transcription works offline; the optional OpenAI Whisper API sends audio to a cloud service.
LocalVocal adds live speech transcription and translation to OBS for streamers and people recording video. It runs Whisper on your own computer, so local captioning works offline and keeps audio processing on your machine. The plugin is open source under GPL-2.0.
OpenWhispr is a free, MIT-licensed dictation and meeting transcription app for people who want voice input across their apps with control over where processing happens. It's available on macOS, Windows, Linux and iOS. Local transcription works offline and keeps audio on your device; optional cloud transcription sends audio to the selected provider, whose retention policies apply.
pyVideoTrans translates spoken audio into another language and produces a video with translated subtitles and AI dubbing. It's for people adapting videos for audiences in other languages who want control over which parts run locally. It recognizes speech directly, so the original video doesn't need subtitles.