AudioWhisper is an open source macOS menu bar app for people who want to dictate into other apps or turn recorded audio into text. It combines offline transcription with optional cloud services, so you can choose where your audio gets processed. The project uses the MIT license.
WhisperKit runs on-device through CoreML, with Whisper models including Tiny, Base, Small and Large Turbo. Parakeet-MLX provides another local option with English and multilingual models. Both work offline once their models are downloaded and don't require API keys. Local Whisper also runs on Intel Macs, though it's slower; Parakeet and local MLX text correction require Apple Silicon. The app requires macOS Sonoma or later.
Global shortcuts, push-to-talk and a windowless recording mode make it suited to dictation while you're working in another app. It copies transcripts to the clipboard and can paste them automatically, then return focus to your previous app. It also transcribes existing audio files. Optional text correction fixes typos and punctuation and removes filler words, with editable categories for contexts such as coding and email.
Correction can run locally through MLX or use a cloud provider. OpenAI Whisper and Google Gemini transcription require internet access and API keys, and send audio to those services. Local transcription keeps audio on your Mac. The app collects no analytics, stores API keys in macOS Keychain, and can keep a searchable local transcript history with retention controls. Its usage dashboard tracks word counts, dictation speed and estimated typing time saved.
Claim this page and we'll verify you by hand. AudioWhisper gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find AudioWhisper?Promote it
Something wrong or outdated on this page?
11.9KUpdated 19 hours agoGPL-3.0
macOS#Multilingual#Streaming inference#Works offline
FluidVoice is a free, open source voice-to-text app for Mac users who want to dictate into email, documents, chat, terminals, and code editors. It transcribes speech locally and can turn rambling dictation into edited text before inserting it into the active app. It requires macOS 15 or later and supports Apple Silicon and Intel Macs, with Whisper providing Intel compatibility.
8.9KUpdated 1 day agoMIT
macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual
1.8KUpdated 9 hours agoGPL-3.0
macOS#Batch processing#Hugging Face integration#MCP
enconvo.comAI Workflow Automation
macOS · iOS#LM Studio integration#MCP#MLX
752Updated 1 week agoGPL-3.0
macOS · Windows · Linux · Docker#LM Studio integration#MLX#Multilingual
1.5KUpdated 7 days agoMIT
macOS · Windows · iOS · Android#Multilingual#Ollama integration#Works offline
OpenWhispr is a free, MIT-licensed dictation and meeting transcription app for people who want voice input across their apps with control over where processing happens. It's available on macOS, Windows, Linux and iOS. Local transcription works offline and keeps audio on your device; optional cloud transcription sends audio to the selected provider, whose retention policies apply.
TypeWhisper is a macOS dictation and transcription app for people who want to keep voice recordings on their own machine and use the resulting text across apps. Local engines keep audio on your Mac; optional cloud providers such as Groq, OpenAI and xAI/Grok process audio remotely.
Enconvo is a native AI assistant and agent for Mac that can use your screen and selected text as context, then work inside your apps. It's for people who want help with writing, research and everyday tasks alongside the app they're using. A sidebar keeps the agent beside the current app, while text selection tools give quick access to editing and translation.
TranscriptionSuite is a local speech-to-text app for people transcribing lectures, recording conversations or dictating into other apps. It's open source under GPL-3.0, with desktop apps for Windows, macOS and Linux. Transcription works offline after the initial app and model downloads, keeping audio and transcripts on your own hardware.
Amical is a free, open-source AI dictation app that formats spoken text for the app you're using. It's for people who want voice input for email, chat, coding prompts and everyday writing, with a choice between local processing and cloud models. It runs on macOS and Windows, with mobile apps for iOS and Android.