91.1KUpdated 22 hours ago
macOS · Windows · Linux#Agent Client Protocol#llama.cpp backend#LM Studio integration
Zed is a native code editor for developers who want AI coding assistance and collaboration in the same desktop app. It runs on macOS, Linux, and Windows, with a minimal interface and a focus on responsive editing. Its AI agents can work in parallel on tasks that involve editing files, exploring code, and running tools.
77KUpdated 23 hours agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image
Unsloth brings model training and everyday AI use into a desktop app for people who want to run models on their own hardware. Its no-code interface covers chat, fine-tuning and media generation on macOS, Windows and Linux. The Unsloth software is open source under Apache 2.0.
116.8KUpdated 5 days agoMIT
#MCP#Ollama integration#Structured output
Browser Use is an MIT licensed browser agent for developers who want AI to carry out tasks on websites. You can run the open source agent on your own machine from Python, choose a model, and use either a local or cloud browser. A CLI is available for browser tasks too.
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
14Updated 20 hours agoMIT
Windows · Linux · Docker · Web#Human approval#LM Studio integration#Multilingual
ScribeDog is a Markdown editor for writers and note-takers who want AI assistance while keeping their documents on their own hardware. It displays formatted text, tables and images instead of Markdown syntax, but saves ordinary .md files that work with Git and other editors. It has a native desktop app and a self-hosted Server Edition available through Docker.
110.7KUpdated 5 hours agoMIT
macOS · Windows · Linux · Docker#Agent Skills#Code execution#MCP
Pi is a terminal-based coding assistant for developers who want to shape the agent around their own workflow. Its core stays small, while extensions can change its tools, commands and terminal interface. It's open source under the MIT license.
43.4KUpdated 1 week agoApache-2.0
macOS · Windows · Linux#Guardrails#Tool calling
agent-browser gives AI agents a compact text view of a browser page, with references that identify the exact elements they can interact with. It's for developers who want coding assistants or other agents to use websites without filling their context with a full page's HTML. The open source tool runs locally on macOS, Linux and Windows under the Apache 2.0 license.
2.2KUpdated 21 hours agoApache-2.0
Docker#LLM tracing#MCP#Multi-user access
Agent Router is an open source AI gateway for teams whose agents use both model APIs and MCP tools. It runs on a laptop, a dedicated gateway, or Kubernetes, and gives applications one OpenAI-compatible entry point for cloud providers and self-hosted inference. The project uses the Apache 2.0 license.
1.4KUpdated 3 weeks ago
Web#Home Assistant integration#OpenAI-compatible API#Tool calling
Extended OpenAI Conversation is a custom component for Home Assistant users who want an AI assistant that can act on their home. It builds on OpenAI Conversation and adds device control, automation creation, and access to historical device states. It runs inside Home Assistant, while a separate model backend handles the conversation.
manual.raycast.comDesktop Chat Apps
#Multimodal input#Ollama integration#Tool calling
Raycast is a proprietary desktop launcher that lets you use models running through Ollama within its AI assistant. It's for people who want private, offline conversations while keeping access to Raycast's chat and AI commands. Local model requests go directly to Ollama on your computer, without sending your conversations to Raycast or third parties. Local model access is paid.
30.4KUpdated 19 hours agoMIT
#MCP#Tool calling
Composio connects AI assistants and custom agents to apps such as Gmail, Slack, GitHub, and Linear. It's for people who want their assistant to act on requests across apps, and developers who don't want to maintain each integration themselves. Its CLI gives coding agents a local interface; the standard setup uses Composio's hosted authentication and execution service. It requires an account and internet access.
3.8KUpdated 2 years agoMIT
Docker#Streaming inference#Tool calling
Vocode is an open source Python library for developers building voice AI agents, with a self-hosted telephony server and support for live conversations through a computer's microphone and speakers. It connects speech recognition, an LLM, and speech synthesis in one library. The code uses the MIT license.
778Updated 2 days agoMIT
Docker#Human approval#Multilingual#Tool calling
Bolna is a voice AI agent platform for businesses handling customer support, lead qualification, reminders, and recruitment calls. Its focus is Indian languages and accents, including Hindi, Hinglish, Tamil, and Telugu, alongside English. Developers can self-host its MIT-licensed Python orchestration core, while business teams can use a hosted dashboard to build agents without code.
90.7KUpdated 1 week ago
Windows#Git integration#Knowledge graphs#MCP
MCP Reference Servers is a collection of locally run examples for developers building connections between AI applications and external tools or data. Maintained by the MCP steering group, the servers demonstrate the protocol and its SDKs. They're educational implementations, so developers should assess security requirements before using them in production.
23.8KUpdated 8 months agoMIT
Web#LLM tracing#Multi-user access#Ollama integration
Vanna is a self-hosted Python framework for building AI agents that answer database questions in plain language. It's for teams adding chat to analytics products or internal data tools, especially when each user needs different access to the data. The project is archived and no longer maintained. Its open-source code uses the MIT license.
5.7KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Ollama integration
OpenAgent is a self-hosted personal AI assistant that combines document search with agents that can act on your behalf. It's for people who want an assistant on their own computer or server, and teams building agents around their documents and workflows. It runs natively on Windows, macOS and Linux as a single executable. It's free and open source under Apache 2.0.
253Updated 4 weeks agoMIT
macOS · Windows#MCP#Multilingual#Multimodal input
Raycast Ollama brings Ollama models into Raycast on macOS and Windows, for people who want AI chat and text assistance within their desktop launcher. It connects to Ollama on your own machine or a remote server you choose. The extension is open source under the MIT license, and local inference doesn't require an Ollama API key.
1.5KUpdated 2 years agoMIT
macOS · Windows · Browser Extension#Multimodal input#Ollama integration#RAG
Lumos is a Chrome extension for asking questions about web pages using models running on your own machine. It's for readers working through long discussions, product reviews, news articles or technical documentation who want summaries and answers tied to the material they're reading. Ollama handles inference locally, without a remote AI server.
6.9KUpdated 1 year agoApache-2.0
Web#Multi-agent workflows#Multilingual#RAG
MindSearch is a self-hosted AI search framework for people who want to build their own Perplexity-style answer engine. Multiple LLM agents search and read web pages to produce answers with a visible research process. It's open source under Apache 2.0.
12.7KUpdated 2 months agoMIT
Windows · Web#Human approval#MCP#Multi-agent workflows
Open Deep Research is a self-hosted AI agent that searches for information and writes research reports. It's for developers and teams who want to choose their own models and research tools. The project is archived and no longer maintained.
11.9KUpdated 2 months agoAGPL-3.0
Web#Code execution#Multilingual#RAG
Scira is an AI search engine for people who want answers backed by sources and control over the application they use for research. It has a hosted website, and you can self-host the open-source application under AGPL-3.0. Its live search relies on internet access and external services; self-hosting doesn't make that research offline.
55.5KUpdated 2 months ago
Docker · Web#Human approval#LLM tracing#Multi-agent workflows
Flowise is a visual builder for AI agents and chatbots that can run locally or on your own server, including through Docker. It's for developers and teams building LLM applications with connected workflow blocks. The project is archived and no longer maintained.
codegpt.coCode Completion and IDE Extensions
VS Code · JetBrains#Code execution#Human approval#LM Studio integration
CodeGPT is an AI coding assistant for VS Code, JetBrains and Visual Studio that lets developers choose between models running locally through Ollama or LM Studio and cloud models such as Claude, GPT, Gemini and Grok. It's for people who want help writing and debugging code inside their editor while keeping control over which model handles each task.
3.5KUpdated 4 months agoBSD-3-Clause
VS Code · JetBrains#Code execution#Persistent memory#RAG
Refact is a self-hosted AI coding assistant for developers who want an agent to work across their repository and development tools. The original project is archived and no longer maintained; ongoing development has moved to JegernOUTT/refact. The original code uses the BSD-3-Clause license.
kerlig.comChat With Your Documents
macOS#LM Studio integration#MCP#Multimodal input
Kerlig is a paid AI writing assistant for Mac that works with text in the apps you already use. It's aimed at people writing client emails, Slack replies and Jira tickets who want help editing and drafting without moving everything into a browser chat. It runs on Apple Silicon and Intel Macs with macOS 12 or later.
35Updated 7 months agoMIT
macOS · Web · VS Code#Code execution#LM Studio integration#MCP
Wingman-AI is a self-hosted AI agent platform for work that needs ongoing context and several agents with different roles. It suits developers and teams handling research, support, or recurring operations alongside coding tasks. A lead agent can delegate work to specialized subagents, each with its own workspace and session history.
4.2KUpdated 1 year agoApache-2.0
Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration
LMQL is a programming language for developers who need model calls and ordinary Python logic in the same program. It lets you define rules for generated text, including types, length limits, allowed answers and stopping phrases. Those rules apply during generation, so you can constrain intermediate responses as well as the final output.
nurgo-software.comAI Workflow Automation
Windows#Code execution#Multi-agent workflows#Multimodal input
BrainSoup is a proprietary native Windows app for people who want custom AI agents to handle work on their desktop. You can give agents distinct roles and access to different data, then have them collaborate in shared chat rooms. Natural-language conversations guide their tasks and automations.
24.3KUpdated 5 months agoApache-2.0
VS Code#MCP#Tool calling
Roo Code is an AI coding assistant for developers working in a code editor. The project is archived and no longer maintained, and its extension has been shut down. It builds on Cline, and its source code is available under the Apache License 2.0.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
14KUpdated 5 months ago
macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP
Open-LLM-VTuber is a local AI companion for people who want a character they can talk to, with a Live2D avatar that responds through speech and expressions. It runs on Windows, macOS and Linux through web and desktop clients. With local models for speech and language processing, it works fully offline and keeps conversations on your device. Cloud APIs are optional alternatives that send the corresponding processing to external services.
10.8KUpdated 3 months agoApache-2.0
#Hugging Face integration#Multilingual#Multimodal input
Voxtral is Mistral AI's open-source audio and text model for developers building self-hosted speech applications. It can answer questions about recordings and produce structured summaries within the same model that transcribes speech. Mistral's separate mistral-inference library is archived and no longer maintained; Voxtral supports vLLM and Hugging Face Transformers.
12.3KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval
ShellGPT is an AI terminal assistant for developers and people who work with shell commands. It runs on Linux, macOS and Windows, turning plain-language requests into commands suited to your operating system and shell. You can review, explain or execute its suggestions, and its Bash and Zsh integrations put generated commands into the terminal input line for editing.
11.7KUpdated 4 months agoApache-2.0
Docker · Web#Batch processing#LLM tracing#Multimodal input
TensorZero is a self-hosted platform for developers building LLM applications. The project is archived and no longer maintained. It combines a model gateway with tools for inspecting responses, evaluating workflows, and improving prompts using production data and human feedback.
typingmind.comChat and Assistants
#Agent Skills#MCP#Multimodal input
TypingMind is a browser chat frontend for people who want to choose their model providers and manage conversations in one place. It connects to cloud APIs such as OpenAI, Claude and Gemini, and supports custom endpoints for locally hosted models such as Ollama and LocalAI. The frontend does not run model inference itself.
4.5KUpdated 7 months agoMIT
macOS · Windows · Linux#MCP#Ollama integration#OpenAI-compatible API
mods is a command-line AI tool for people who want to ask questions about command output or use model responses in shell pipelines. It connects to local LLMs through LocalAI as well as cloud services. The project is archived and no longer maintained.