4.9KUpdated 7 months agoApache-2.0
macOS#LoRA#Multilingual
ACE-Step is an open-source music generation model for musicians, producers and developers who want to create and edit music on their own hardware. It generates songs with vocals or instrumental tracks from text descriptions and supplied lyrics. You can choose the duration and describe the sound with genre tags, longer prompts or a scene description.
77KUpdated 3 months agoAGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#MCP#Multi-user access#RAG
AppFlowy is a self-hosted Notion alternative for teams that want project management, shared documents, and AI in an environment they control. Its self-hosted enterprise offering can run on premises, in your own cloud, or in an air-gapped environment. Self-hosted LLMs and embedding models let you keep AI processing and workspace data inside your infrastructure, with no vendor access to your instance.
26.4KUpdated 2 years agoMIT
macOS · Windows · Linux
Ultimate Vocal Remover is a desktop app for separating vocals and other stems from audio files on your own computer. It's for people making karaoke backing tracks, isolating a vocal, or working with separate parts of a song. It runs on Windows, macOS and Linux, and its MIT license allows you to use and modify the software.
25.6KUpdated 3 hours agoMIT
#Batch processing#Hugging Face integration#Quantization
faster-whisper is a Python library for people building local speech transcription into their own software. It runs OpenAI's Whisper models through CTranslate2, with faster processing and lower memory use than the original Whisper implementation in the project's comparisons. It runs on a CPU. NVIDIA GPUs are supported too, and the code is open source under the MIT license.
109.8KUpdated 4 weeks agoMIT
#Multilingual#Voice activity detection
Whisper is an open source speech recognition model for people who want to transcribe audio on their own hardware. It suits developers adding voice features to an app and anyone working with recordings in multiple languages. OpenAI publishes the models and inference code under the MIT license, so audio can stay on the machine running them.
24.8KUpdated 21 hours agoApache-2.0
iOS · Android#Agent Skills#Hugging Face integration#Multilingual
Google AI Edge Gallery is an open-source app for people who want to try generative AI on their own phone. It runs model inference locally, so offline chat, image analysis and audio tasks don't send your inputs to a server. It supports Android and iOS and uses the Apache 2.0 license.
apps.apple.comDictation and Voice Typing
iOS#Multilingual#Works offline
Google AI Edge Eloquent is an AI dictation app for iPhone and iPad that turns spoken thoughts into edited text using on-device Gemma models. It's for people who prefer speaking to typing but don't want every hesitation or self-correction carried into their notes and drafts.