91.1KUpdated 21 hours ago
macOS · Windows · Linux#Agent Client Protocol#llama.cpp backend#LM Studio integration
Zed is a native code editor for developers who want AI coding assistance and collaboration in the same desktop app. It runs on macOS, Linux, and Windows, with a minimal interface and a focus on responsive editing. Its AI agents can work in parallel on tasks that involve editing files, exploring code, and running tools.
77KUpdated 22 hours agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image
Unsloth brings model training and everyday AI use into a desktop app for people who want to run models on their own hardware. Its no-code interface covers chat, fine-tuning and media generation on macOS, Windows and Linux. The Unsloth software is open source under Apache 2.0.
5.7KUpdated 2 days agoGPL-3.0
#ONNX
Piper turns text into spoken audio on local hardware. It's an open-source, GPL-3.0 text-to-speech engine for developers building voice features, accessibility tools and self-hosted AI projects. Speech generation runs locally, giving people a way to add a voice to software they control.
116.8KUpdated 5 days agoMIT
#MCP#Ollama integration#Structured output
Browser Use is an MIT licensed browser agent for developers who want AI to carry out tasks on websites. You can run the open source agent on your own machine from Python, choose a model, and use either a local or cloud browser. A CLI is available for browser tasks too.
15KUpdated 11 months ago
macOS · Windows · Linux · Web#Hugging Face integration
Hunyuan3D generates textured 3D assets from reference images or text and can run on your own computer. It's for 3D artists, hobbyists and developers who want AI-generated meshes they can use in other software. It supports macOS, Windows and Linux; Hunyuan3D Studio is a separate hosted option.
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
whispernotes.appChat With Your Documents
macOS · iOS#MCP#Multilingual#Speaker diarization
Whisper Notes is an offline transcription app for iPhone, iPad and Apple Silicon Macs. It's for people recording interviews, lectures or meetings who need the audio and transcripts to stay on their device. It requires no account and has no cloud sync, analytics or tracking.
picapport.deAI Photo Libraries
macOS · Windows · Linux · Docker · Web#Batch processing#Multi-user access#Role-based access
PicApport is a self-hosted photo server for families and businesses that want a searchable media archive on their own hardware. Its AI tagging add-on recognizes image content without sending photos to a cloud service and saves the resulting tags in image metadata. The server runs on Windows, Linux or macOS, with a Docker image available, and you access the gallery through a browser.
14Updated 20 hours agoMIT
Windows · Linux · Docker · Web#Human approval#LM Studio integration#Multilingual
ScribeDog is a Markdown editor for writers and note-takers who want AI assistance while keeping their documents on their own hardware. It displays formatted text, tables and images instead of Markdown syntax, but saves ordinary .md files that work with Git and other editors. It has a native desktop app and a self-hosted Server Edition available through Docker.
110.7KUpdated 4 hours agoMIT
macOS · Windows · Linux · Docker#Agent Skills#Code execution#MCP
Pi is a terminal-based coding assistant for developers who want to shape the agent around their own workflow. Its core stays small, while extensions can change its tools, commands and terminal interface. It's open source under the MIT license.
43.4KUpdated 1 week agoApache-2.0
macOS · Windows · Linux#Guardrails#Tool calling
agent-browser gives AI agents a compact text view of a browser page, with references that identify the exact elements they can interact with. It's for developers who want coding assistants or other agents to use websites without filling their context with a full page's HTML. The open source tool runs locally on macOS, Linux and Windows under the Apache 2.0 license.
30.1KUpdated 3 weeks agoMIT
macOS#GGUF#Hugging Face integration#Hybrid search
QMD is a local search engine for people with Markdown notes, meeting transcripts, or documentation they want to search themselves or make available to an AI agent. It accepts exact keywords and natural-language queries, with indexing and model inference running on your own machine. It's open source under MIT.
8.9KUpdated 10 hours agoMIT
macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual
OpenWhispr is a free, MIT-licensed dictation and meeting transcription app for people who want voice input across their apps with control over where processing happens. It's available on macOS, Windows, Linux and iOS. Local transcription works offline and keeps audio on your device; optional cloud transcription sends audio to the selected provider, whose retention policies apply.
1.5KUpdated 5 days agoMIT
macOS · Windows · iOS · Android#Multilingual#Ollama integration#Works offline
Amical is a free, open-source AI dictation app that formats spoken text for the app you're using. It's for people who want voice input for email, chat, coding prompts and everyday writing, with a choice between local processing and cloud models. It runs on macOS and Windows, with mobile apps for iOS and Android.
6.5KUpdated 6 days agoAGPL-3.0
Linux · Docker · Web#Multi-user access
Photoview is a self-hosted photo gallery for photographers and households who keep their pictures on a personal server or NAS. It turns existing folders into browser-based albums and automatically scans for added photos and videos. Your folder structure determines how the library appears, so you can keep organizing files through Samba, FTP or Nextcloud. Your media stays on your server.
locallyai.appDesktop Chat Apps
macOS · iOS#MLX#Multilingual#Multimodal input
Locally AI is a native app for running language and vision models on recent iPhones, iPads, and Macs. It's for people who want a private AI assistant on their own device, with text, image processing, and voice conversations available without cloud processing. Once a model is downloaded, it works offline and doesn't require an account.
17.1KUpdated 2 weeks ago
Windows#Batch processing#Works offline
Waifu2x-Extension-GUI is a local AI upscaler for Windows users working with anime, photos and video. It combines image enlargement, noise reduction and video frame interpolation in a desktop interface. Media processing stays on your PC; the app doesn't upload your files or collect user data.
recurse.chatChat With Your Documents
macOS#GGUF#Hugging Face integration#OpenAI-compatible API
RecurseChat is a paid Mac app for people who want to chat with AI and ask questions about their files on their own computer. It runs local LLMs without an internet connection or a server, keeping local conversations on the device. The same app also connects to Claude and ChatGPT, so you can choose local processing or a cloud provider.
sindresorhus.comOn-Device and In-Browser AI
macOS · iOS#Batch processing#Multilingual
Aiko is a paid, native transcription app for macOS, iOS and visionOS that processes speech on your device with OpenAI's Whisper model. It's for people turning meetings, lectures or other recordings into text while keeping the audio local, including sensitive recordings.
1.1KUpdated 4 weeks agoBSD-3-Clause
Android · Browser Extension#Multilingual#Ollama integration#Works offline
Linguist is a free browser translation extension for people who read across languages and want control over where their text goes. Its built-in Bergamot translator processes text on your device without an internet connection. You can also choose an external translation provider or connect a local backend such as Ollama or LibreTranslate.
12KUpdated 1 week agoApache-2.0
Docker · Web#Batch processing#Human approval#Multi-agent workflows
Bisheng is an open source, self-hosted platform for teams building AI applications around business documents and processes. Its visual workflow editor combines automated tasks with human feedback, including intervention during multi-turn conversations. It's suited to document review, support ticket assistance and report generation that need more control than a single chatbot exchange.
prodi.gyData Labeling and Annotation
Web#Hugging Face integration#Works offline
Prodigy is a proprietary annotation tool that runs on your own machines, including air-gapped systems without an internet connection. It's for developers and research teams building training and evaluation datasets for custom AI models. The Python library includes a web application where annotators can label data without programming knowledge.
3.3KUpdated 1 month agoApache-2.0
Docker · Web#Multi-user access#Prompt versioning
Pezzo is an open-source platform for developers and teams managing prompts and monitoring LLM applications. You can run the full stack locally with Docker Compose, keeping the prompt management and monitoring platform on infrastructure you control. Its source code uses the Apache 2.0 license.
perfectmemory.aiAI Notes and Knowledge Bases
macOS · Windows#LM Studio integration#MCP#Ollama integration
Perfect Memory AI keeps a searchable record of what you've seen on your screen and heard in meetings on Mac and Windows. It's for people who need to find a page, presentation or conversation again without remembering which app held it. Recordings stay on your device, and search returns material you've actually seen or heard.
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.
3.2KUpdated 3 months agoMIT
#Guardrails
LLM Guard is a Python security toolkit for developers building applications around large language models. It checks prompts and generated responses for risks such as prompt injection, sensitive data exposure and harmful language. The project is archived and no longer maintained, including its associated models on Hugging Face.
4.6KUpdated 19 hours agoMIT
Web#AI red teaming#OpenAI-compatible API
PyRIT is an MIT-licensed, open source Python framework for security professionals and engineers assessing generative AI systems. It combines automated attack testing with human-led investigations through CoPyRIT, a web interface served locally. The framework runs locally, but prompts go to the target services you choose; cloud targets and cloud-based scorers process requests outside your machine.
1.5KUpdated 3 years agoApache-2.0
Web#Guardrails#Semantic search
Rebuff is a prompt injection detector for developers building LLM applications that accept untrusted input. It combines checks for suspicious prompts with a record of past attacks and tests for leaked prompt content. The project is archived and no longer maintained.
496Updated 3 years agoApache-2.0
Docker · Web#Guardrails#Semantic search
Vigil is a self-hosted security scanner for developers and researchers who want to check LLM inputs and responses for prompt injection, jailbreak attempts, and other suspicious content. It combines several detection methods and includes attack signatures and datasets, so teams can assess known threats without building every detector themselves. It is experimental alpha software for research and is open source under Apache 2.0.
1.2KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing
GoModel is a self-hosted AI gateway for developers and platform teams that want one API for local models and cloud providers. It accepts OpenAI- and Anthropic-compatible requests, so applications can keep their existing SDKs while the gateway handles provider selection and usage controls.
2.2KUpdated 20 hours agoApache-2.0
Docker#LLM tracing#MCP#Multi-user access
Agent Router is an open source AI gateway for teams whose agents use both model APIs and MCP tools. It runs on a laptop, a dedicated gateway, or Kubernetes, and gives applications one OpenAI-compatible entry point for cloud providers and self-hosted inference. The project uses the Apache 2.0 license.
1.2KUpdated 2 years agoMIT
Docker#Guardrails#Multi-user access#OpenAI-compatible API
BricksLLM is a self-hosted AI gateway for teams that need to control access and spending across LLM applications. It sits between applications and model providers, applying cost and rate limits to individual API keys. The gateway is open source under the MIT license and runs locally or on your server through Docker, with PostgreSQL and Redis.
5.6KUpdated 2 years agoApache-2.0
#Ollama integration#OpenAI-compatible API
RouteLLM is a self-hosted Python framework for developers who want to split requests between a stronger LLM and a cheaper model. It judges which prompts need the stronger model, so an application doesn't have to send every request to its most expensive provider. You control the cost-quality tradeoff through a routing threshold.
37KUpdated 2 years agoMIT
Linux · Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API
One API is a self-hosted LLM API gateway for developers and teams that want to share model access across apps or users. It puts cloud providers and Ollama behind an OpenAI-compatible API, so clients can use one endpoint across different backends. It's open source under MIT and runs on your own server as a single executable or in Docker.
1.4KUpdated 3 weeks ago
Web#Home Assistant integration#OpenAI-compatible API#Tool calling
Extended OpenAI Conversation is a custom component for Home Assistant users who want an AI assistant that can act on their home. It builds on OpenAI Conversation and adds device control, automation creation, and access to historical device states. It runs inside Home Assistant, while a separate model backend handles the conversation.
318Updated 3 weeks agoApache-2.0
macOS · Windows · Linux · iOS · Android · Web#Quantization
picoLLM is an on-device inference SDK for developers building apps that run compressed language models on users' hardware. It generates text locally, so prompts don't need to go to a cloud inference service. Its main distinction is Picovoice's compression method, which learns how to allocate precision across model weights rather than applying a fixed allocation.
223Updated 17 hours agoApache-2.0
macOS · Linux#GGUF#Git integration#Guardrails
LLMKube is a free, open-source Kubernetes operator for teams and homelab owners running local LLM inference across their own hardware. It manages Linux GPU servers and Apple Silicon Macs together, so a mixed fleet can serve models through the same platform. It uses the Apache 2.0 license.
592Updated 5 days agoMIT
Docker#Ollama integration
Ollama Helm Chart packages Ollama for teams that want to run a local LLM service on their own Kubernetes cluster. It's a community-maintained, open source chart under the MIT license, aimed at developers and infrastructure teams managing AI alongside other cluster services.
1.2KUpdated 8 months agoMIT
Linux · Docker#Code execution#Home Assistant integration#Voice activity detection
Wyoming Satellite connects a microphone and audio playback device to Home Assistant through the Wyoming protocol. It's for people building a self-hosted voice assistant with a separate device for speaking and listening. The project is archived and no longer maintained; its replacement is Linux Voice Assistant, which uses the ESPHome protocol.
28.5KUpdated 3 months agoMIT
Windows · Docker
Spleeter is Deezer's music source separation library for developers and audio researchers who want to split recordings into separate vocal and instrumental tracks on their own hardware. It includes pretrained models, so you can separate audio without first training a model. The library is open source under the MIT license.
1.1KUpdated 2 years agoMIT
#Hugging Face integration#LoRA#Quantization
DataDreamer connects LLM prompting, synthetic data generation, and model training in one Python library. It's for researchers and developers who want to build datasets and use them to fine-tune or align models in reproducible workflows. The library is open source under the MIT license.
manual.raycast.comDesktop Chat Apps
#Multimodal input#Ollama integration#Tool calling
Raycast is a proprietary desktop launcher that lets you use models running through Ollama within its AI assistant. It's for people who want private, offline conversations while keeping access to Raycast's chat and AI commands. Local model requests go directly to Ollama on your computer, without sending your conversations to Raycast or third parties. Local model access is paid.
30.4KUpdated 18 hours agoMIT
#MCP#Tool calling
Composio connects AI assistants and custom agents to apps such as Gmail, Slack, GitHub, and Linear. It's for people who want their assistant to act on requests across apps, and developers who don't want to maintain each integration themselves. Its CLI gives coding agents a local interface; the standard setup uses Composio's hosted authentication and execution service. It requires an account and internet access.
6.8KUpdated 4 weeks agoMIT
Windows · Linux
ROCm is AMD's open-source GPU computing platform for developers running AI training, inference and scientific workloads on their own hardware or servers. It supports selected Linux and Windows configurations on AMD Instinct, Radeon and Ryzen AI devices. Check the version-specific GPU, operating-system, driver and firmware compatibility matrix before installing. It's the software foundation for applications that need AMD GPU acceleration, including local LLM workloads.
4.6KUpdated 4 years agoMIT
macOS
asitop is a terminal hardware monitor for Apple Silicon Macs, with separate views of CPU, GPU and Apple Neural Engine activity. Its README targets Apple Silicon on macOS Monterey; hardware-counter availability can differ on newer macOS releases. It runs locally and suits people who want to watch hardware use during demanding workloads, including local AI inference. It's free and open source under the MIT license.
10.4KUpdated 3 years agoMIT
macOS · Windows · Linux · Docker#Batch processing#Quantization
Demucs separates a finished song into vocals, drums, bass and the remaining accompaniment on your own computer. It's for musicians who need individual stems or a vocal-free backing track, and developers building audio tools. The project is archived and no longer maintained. Its Python code is open source under the MIT license.
3.8KUpdated 2 years agoMIT
Docker#Streaming inference#Tool calling
Vocode is an open source Python library for developers building voice AI agents, with a self-hosted telephony server and support for live conversations through a computer's microphone and speakers. It connects speech recognition, an LLM, and speech synthesis in one library. The code uses the MIT license.
10.8KUpdated 2 days agoMIT
Browser Extension#Multilingual#Ollama integration#OpenAI-compatible API
ChatGPTBox is a free, open-source browser extension for people who want AI help with the pages they read, search results and selected text. It works in Chrome, Edge, Firefox and Safari, with mobile support too. You can connect it to Ollama or a self-hosted model through its custom model mode, or use cloud services such as ChatGPT, Claude, Moonshot and Azure.