Local AI for Speech, Voice and Music

Speech recognition, dictation, text-to-speech, voice cloning and music generation that run on your computer rather than through a paid API.

Subcategories

100+ tools
Favicon of ACE-Step

ACE-Step

3 videos
An open-source AI music model that runs on NVIDIA GPUs and Apple Silicon, with text-to-music generation, adjustable duration and localized lyric editing.

4.9KUpdated 7 months agoApache-2.0

macOS#LoRA#Multilingual

Favicon of AppFlowy

AppFlowy

1 video
A self-hosted AI workspace for projects and wikis, with local LLM support, air-gapped deployment, and apps for desktop, mobile, and the web.

77KUpdated 3 months agoAGPL-3.0

macOS · Windows · Linux · iOS · Android · Web#MCP#Multi-user access#RAG

An open-source desktop vocal remover that processes audio locally on Windows, macOS and Linux using neural source separation models.

26.4KUpdated 2 years agoMIT

macOS · Windows · Linux

An open-source Python library for local speech transcription with Whisper models. It runs on CPUs or NVIDIA GPUs and uses CTranslate2.

25.6KUpdated 3 hours agoMIT

#Batch processing#Hugging Face integration#Quantization

Favicon of Whisper

Whisper

6 videos
An MIT-licensed speech recognition model that runs on your own hardware, transcribes multiple languages and translates speech into English.

109.8KUpdated 4 weeks agoMIT

#Multilingual#Voice activity detection

Local AI app for Android and iOS. Run Gemma 4 on-device, ask questions about photos, transcribe audio and compare model performance.

24.8KUpdated 21 hours agoApache-2.0

iOS · Android#Agent Skills#Hugging Face integration#Multilingual

An AI dictation app for iPhone and iPad that uses Gemma models locally. It supports offline transcription and optional cloud text features.

apps.apple.comDictation and Voice Typing

iOS#Multilingual#Works offline