407Updated 2 days agoMIT
macOS#MLX#Multimodal input#OpenAI-compatible API
Slotstream runs Qwen3.8-Flash-Next on Apple Silicon Macs that don't have enough RAM to hold the whole model. It's aimed at people with 16 to 64 GB of memory who want local chat, image questions or a model backend for coding agents. Most model weights stay on the SSD, while frequently used expert networks stay in memory. The full model remains available.
32.5KUpdated 3 days agoMIT
macOS · Windows · Linux#GGUF#Hugging Face integration#Voice activity detection
Handy is a free, MIT-licensed speech-to-text app for people who want to dictate wherever they type on a computer. It runs on Windows, macOS and Linux. Transcription happens locally, so your voice stays on your machine and the app can work offline.
31.3KUpdated 3 weeks agoMIT
macOS · Windows · Linux#Ollama integration#OpenAI-compatible API#Streaming inference
Meetily is a local AI meeting assistant for people who want meeting notes while keeping recordings on their own device. It captures calls from Zoom, Google Meet, Microsoft Teams and other meeting software without placing a bot in the meeting. You can watch the transcript appear during the call, then generate a summary.
742Updated 4 hours agoAGPL-3.0
Web#MCP#Tool calling#Works offline
CozyClay is a local 3D previsualization studio for filmmakers and AI video creators who want to plan framing and movement before generating a finished clip. You build a rough scene, pose characters, and edit camera moves and cuts on a timeline. Those shots become visual references for models such as Seedance, Kling and Veo, so you can specify the camera through a clip rather than relying on text alone.
2.4KUpdated 1 day agoApache-2.0
macOS · Windows · Linux#Agent Skills#Git integration#MCP
ripwire helps AI coding agents find relevant code and assess changes without filling their context with whole files. It runs offline on your machine as a self-contained binary, with a CLI and an optional MCP server. It's for developers who want their agents to spend less context on repository research and check the effects of their edits.
1.1KUpdated 8 hours agoMIT
macOS · Windows · Linux#Agent Skills#MCP#Persistent memory
deja-vu gives coding assistants a shared memory of work already recorded on your machine. It searches sessions from before you installed it, so developers can recover an old fix or carry context between Claude Code, Codex CLI, Cursor and opencode without starting a separate collection of notes.
16.3KUpdated 2 days ago
macOS · Windows · Linux#LM Studio integration#Ollama integration#OpenAI-compatible API
SurfSense is an open-source NotebookLM alternative that turns documents into cited answers and editable deliverables on your computer. It's a desktop app for Windows, macOS and Linux, aimed at people working with confidential files, research or study material. You don't need an account.
6.6KUpdated 1 day ago
macOS · iOS#Works offline
VoiceInk is a native macOS dictation app for people who want to write emails, notes, and AI prompts by speaking. It transcribes speech locally and works across applications, so writers, students, and developers can dictate into the apps they already use. Voice transcription works offline.
47.7KUpdated 1 month agoApache-2.0
macOS · Linux · Web#Distributed execution#Hugging Face integration#MLX
exo is a local LLM runner that combines your devices into a cluster, letting you use models too large for one machine's memory. It's for people who want to run large models on their own hardware and developers connecting existing AI clients to local inference. It runs on macOS and Linux under the Apache 2.0 license.
17KUpdated 3 days agoAGPL-3.0
Web#Multilingual#Works offline
LibreTranslate is a self-hosted machine translation API for developers and organizations that want to run translation on their own servers. It's free to download and can work offline. The project is open source under the GNU Affero General Public License v3.0 (AGPL-3.0).
39.9KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#Knowledge graphs#LLM tracing#Multimodal input
LightRAG combines knowledge graphs with vector search to answer questions across a document collection. It's a self-hosted Python framework for developers building document assistants, particularly where answers depend on relationships between facts in different files, such as legal or financial material.
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
lmstudio.aiComputer and Browser Agents
macOS · Windows · Linux#llama.cpp backend#MCP#MLX
LM Studio is a desktop application for downloading and running language models on macOS, Windows and Linux. You can search for models, manage downloads and chat with them through the app. Downloaded models can run offline, including document chat that uses files on your computer.
36.2KUpdated 7 hours agoMIT
#Home Assistant integration#Works offline
Frigate is an open source NVR for IP cameras that detects people, cars, and other objects on your own hardware. Camera feeds stay at home. It's built for people who want a locally controlled security camera system, particularly those using Home Assistant.
26.6KUpdated 2 months agoMIT
Linux#Multilingual#Voice cloning#Voice conversion
Chatterbox is an MIT-licensed text-to-speech model family for developers and creators who want to generate speech on their own hardware. You can self-host it on a GPU, including in an air-gapped environment, without an account or API key. Resemble AI also offers separate managed hosting.
5.8KUpdated 11 hours agoApache-2.0
Android#LM Studio integration#LoRA#Multilingual
Gemma is Google DeepMind’s family of open-weight AI models for developers building applications that can run on their own hardware. Its range covers compact models for phones and IoT devices alongside larger Gemma 4 models for reasoning on personal computers and servers. Some applications can work offline, keeping model inference on the device. Google AI Studio and Google Cloud are also available for hosted use.
135.6KUpdated 1 hour agoGPL-3.0
macOS · Windows · Linux · Web#ControlNet#Inpainting#LoRA
ComfyUI is a local visual AI workspace for artists and technical teams who want to control how images, video, audio, 3D models and text are made. Its node canvas shows each model and processing step, so users can build and adjust workflows without writing code. It runs on your hardware.
52.3KUpdated 1 hour agoAGPL-3.0
macOS · Windows · Linux#LM Studio integration#MCP#Multilingual
Cherry Studio is a free, open source desktop app for people who use several AI models and want their conversations in one place. It runs on Windows, macOS and Linux. Chats, settings and knowledge base files are stored on your device.
54KUpdated 2 days agoMIT
macOS · Windows · Linux · iOS · Android · Docker#Hugging Face integration#Quantization#Streaming inference
whisper.cpp runs OpenAI's Whisper speech recognition models on your own hardware, with fully offline transcription once you've downloaded a model. It's for developers building speech-to-text into applications and people who want to transcribe audio locally. Audio can stay on-device rather than going to a cloud transcription service. The project is open source under the MIT license.
68.2KUpdated 1 day agoMIT
macOS · Windows · Linux#MCP#Works offline
Docling is an MIT-licensed, open source document parser for developers turning files into structured content for search and AI applications. It runs locally on macOS, Linux, and Windows, including in air-gapped environments. Its PDF processing identifies page layout and reading order, extracts tables, code, and formulas, and classifies images.
153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
whispernotes.appChat With Your Documents
macOS · iOS#MCP#Multilingual#Speaker diarization
Whisper Notes is an offline transcription app for iPhone, iPad and Apple Silicon Macs. It's for people recording interviews, lectures or meetings who need the audio and transcripts to stay on their device. It requires no account and has no cloud sync, analytics or tracking.
14Updated 20 hours agoMIT
Windows · Linux · Docker · Web#Human approval#LM Studio integration#Multilingual
ScribeDog is a Markdown editor for writers and note-takers who want AI assistance while keeping their documents on their own hardware. It displays formatted text, tables and images instead of Markdown syntax, but saves ordinary .md files that work with Git and other editors. It has a native desktop app and a self-hosted Server Edition available through Docker.
8.9KUpdated 10 hours agoMIT
macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual
OpenWhispr is a free, MIT-licensed dictation and meeting transcription app for people who want voice input across their apps with control over where processing happens. It's available on macOS, Windows, Linux and iOS. Local transcription works offline and keeps audio on your device; optional cloud transcription sends audio to the selected provider, whose retention policies apply.
1.5KUpdated 5 days agoMIT
macOS · Windows · iOS · Android#Multilingual#Ollama integration#Works offline
Amical is a free, open-source AI dictation app that formats spoken text for the app you're using. It's for people who want voice input for email, chat, coding prompts and everyday writing, with a choice between local processing and cloud models. It runs on macOS and Windows, with mobile apps for iOS and Android.
locallyai.appDesktop Chat Apps
macOS · iOS#MLX#Multilingual#Multimodal input
Locally AI is a native app for running language and vision models on recent iPhones, iPads, and Macs. It's for people who want a private AI assistant on their own device, with text, image processing, and voice conversations available without cloud processing. Once a model is downloaded, it works offline and doesn't require an account.
17.1KUpdated 2 weeks ago
Windows#Batch processing#Works offline
Waifu2x-Extension-GUI is a local AI upscaler for Windows users working with anime, photos and video. It combines image enlargement, noise reduction and video frame interpolation in a desktop interface. Media processing stays on your PC; the app doesn't upload your files or collect user data.
recurse.chatChat With Your Documents
macOS#GGUF#Hugging Face integration#OpenAI-compatible API
RecurseChat is a paid Mac app for people who want to chat with AI and ask questions about their files on their own computer. It runs local LLMs without an internet connection or a server, keeping local conversations on the device. The same app also connects to Claude and ChatGPT, so you can choose local processing or a cloud provider.
1.1KUpdated 4 weeks agoBSD-3-Clause
Android · Browser Extension#Multilingual#Ollama integration#Works offline
Linguist is a free browser translation extension for people who read across languages and want control over where their text goes. Its built-in Bergamot translator processes text on your device without an internet connection. You can also choose an external translation provider or connect a local backend such as Ollama or LibreTranslate.
prodi.gyData Labeling and Annotation
Web#Hugging Face integration#Works offline
Prodigy is a proprietary annotation tool that runs on your own machines, including air-gapped systems without an internet connection. It's for developers and research teams building training and evaluation datasets for custom AI models. The Python library includes a web application where annotators can label data without programming knowledge.
perfectmemory.aiAI Notes and Knowledge Bases
macOS · Windows#LM Studio integration#MCP#Ollama integration
Perfect Memory AI keeps a searchable record of what you've seen on your screen and heard in meetings on Mac and Windows. It's for people who need to find a page, presentation or conversation again without remembering which app held it. Recordings stay on your device, and search returns material you've actually seen or heard.
manual.raycast.comDesktop Chat Apps
#Multimodal input#Ollama integration#Tool calling
Raycast is a proprietary desktop launcher that lets you use models running through Ollama within its AI assistant. It's for people who want private, offline conversations while keeping access to Raycast's chat and AI commands. Local model requests go directly to Ollama on your computer, without sending your conversations to Raycast or third parties. Local model access is paid.
28.1KUpdated 3 years agoAGPL-3.0
#Hugging Face integration#ONNX#Voice conversion
so-vits-svc is an offline AI framework for changing the voice in an existing singing recording while preserving its pitch and intonation. It's aimed at developers and researchers who want to train their own singing voices, including fictional character voices. The project is archived and no longer maintained.
818Updated 4 years agoApache-2.0
macOS · Windows · Linux · Docker#Home Assistant integration#Works offline
DeepStack is a self-hosted computer vision API for developers adding image analysis to camera systems, home automation or other applications. It runs prebuilt and custom models on your own hardware and works fully offline. Image processing stays on the device or server where you host it, with no cloud service required.
982Updated 1 year ago
macOS · Windows · Linux · Docker#Home Assistant integration#Image-to-image#Multimodal input
CodeProject.AI Server gives developers a shared API for AI tasks that run on their own hardware. It's a self-hosted service for adding image analysis, text processing and generation to applications. Processing stays on the machine running the server, without cloud calls or sending data outside your device or network.