8.5KUpdated 23 hours agoApache-2.0
Web#LLM tracing#MCP#Multimodal input
Bifrost is a self-hosted AI gateway for developers and teams whose applications use multiple model providers. It puts Ollama, custom model deployments, and cloud services behind one OpenAI-compatible API, so applications can switch models without maintaining a separate integration for each provider.
186.5KUpdated 1 day agoAGPL-3.0
Docker#Batch processing#MCP#Structured output
Firecrawl helps developers give AI agents and applications access to current web content. It searches for pages, extracts their contents, and returns data in forms an application can use. The project is open source under AGPL-3.0 and can run on a self-hosted server. Firecrawl also offers a hosted service that requires an account and API key and includes additional features. Reaching live websites requires an internet connection.
32.3KUpdated 1 day ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
29.8KUpdated 4 days agoApache-2.0
Docker · Web#MCP#Multi-agent workflows#Ollama integration
GPT Researcher is a self-hosted AI research agent for people who need detailed, cited reports drawn from the web, their own documents, or both. You choose the language model and search provider, and can run the software on your own server or through Docker. It's open source under Apache 2.0.
45.2KUpdated 5 hours agoMIT
Web#Agent Skills#Code execution#MCP
LibreChat is a self-hosted ChatGPT alternative for people and teams that want local models and cloud AI services in one chat interface. It runs on a server you control and is open source under the MIT license. You can switch models without moving to a separate chat app.
37.8KUpdated 1 day agoAGPL-3.0
Docker · Web#Web search
SearXNG is a self-hosted metasearch engine for people who want web search without user tracking or profiling. It brings results from separate search services into one browser interface. You can run your own instance or use a public one, depending on who you want to trust with your searches.
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
20.3KUpdated 2 hours agoMIT
#Human approval#LLM tracing#MCP
Pydantic AI is a Python SDK for developers building AI agents into their own applications. Its main draw is Pydantic validation across agent tools and results, so an agent can return structured data that application code can check and use. The SDK is MIT licensed.
127.4KUpdated 12 minutes agoApache-2.0
macOS · Windows · Linux#Code execution#Git integration#Human approval
Codex CLI is an open source coding assistant that runs on your computer and works with local repositories. It's for developers who want to explore code, make changes, and use their existing development tools from the terminal. It runs on macOS, Windows, and Linux under the Apache 2.0 license.
5.7KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Ollama integration
OpenAgent is a self-hosted personal AI assistant that combines document search with agents that can act on your behalf. It's for people who want an assistant on their own computer or server, and teams building agents around their documents and workflows. It runs natively on Windows, macOS and Linux as a single executable. It's free and open source under Apache 2.0.
14.2KUpdated 1 week agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
QAnything is a self-hosted knowledge base for people and teams who want to ask questions about their own documents, including collections that mix Chinese and English. It can answer in either language regardless of the document's language, and runs locally through Docker on Windows, macOS and Linux.
6.9KUpdated 1 year agoApache-2.0
Web#Multi-agent workflows#Multilingual#RAG
MindSearch is a self-hosted AI search framework for people who want to build their own Perplexity-style answer engine. Multiple LLM agents search and read web pages to produce answers with a visible research process. It's open source under Apache 2.0.
12.7KUpdated 2 months agoMIT
Windows · Web#Human approval#MCP#Multi-agent workflows
Open Deep Research is a self-hosted AI agent that searches for information and writes research reports. It's for developers and teams who want to choose their own models and research tools. The project is archived and no longer maintained.
11.9KUpdated 2 months agoAGPL-3.0
Web#Code execution#Multilingual#RAG
Scira is an AI search engine for people who want answers backed by sources and control over the application they use for research. It has a hosted website, and you can self-host the open-source application under AGPL-3.0. Its live search relies on internet access and external services; self-hosting doesn't make that research offline.
31.5KUpdated 1 year agoMIT
Web#Multi-agent workflows#RAG#Source citations
STORM is an MIT-licensed, open-source research tool from Stanford OVAL that turns a topic into a Wikipedia-style report with citations. You can run its Python software on your own machine or host it on a server; Stanford also offers a hosted research preview. It's aimed at people gathering background material and writers preparing an article draft.
9.2KUpdated 5 days agoApache-2.0
Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API
Morphic is a self-hosted AI search engine for people who want answers backed by web sources and control over the search interface they use. It combines web searches and URL reading with AI-generated responses, so you can examine the sources behind an answer. It's open source under Apache 2.0, and you can run your own instance with Docker or use the hosted website.
3.5KUpdated 2 years agoApache-2.0
Docker · Web#Ollama integration#Web search
Farfalle is a self-hosted AI search engine for people who want a Perplexity-style search app with a choice of local or cloud models. It combines web search with model-generated answers and includes an agent that plans and carries out searches. The code is open source under Apache 2.0.
19.6KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker · Web#Ollama integration#Web search
Devika is a self-hosted AI coding agent for developers who want to give a software task in plain language and have an agent plan the work, research it and write code. Modeled after Cognition AI's Devin, it runs on your own machine with a browser interface and supports local LLMs through Ollama. It's open source under the MIT license.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
mindmac.appDesktop Chat Apps
macOS#llama.cpp backend#LM Studio integration#MLX
MindMac is a native macOS AI chat client for people who want to use local LLMs and cloud services in the same app. Its inline mode lets you ask questions or generate text inside Notes, Mail and browsers without switching to a chat window. It runs on Intel and Apple Silicon Macs with macOS 13 or newer.
39.6KUpdated 1 year ago
#Ollama integration#RAG#Reranking
Quivr Core is a Python framework for developers adding document-based AI answers to their own applications. It combines file ingestion with retrieval-augmented generation (RAG), so a model can answer questions using material from your documents. It supports local models through Ollama as well as cloud APIs from OpenAI, Anthropic, and Mistral.
8.4KUpdated 21 hours agoMIT
#Agent Skills#Batch processing#Guardrails
OGX, formerly Llama Stack, is a self-hosted AI application server for developers building chat apps, document search or AI agents. It brings model inference, file storage, vector search and agent orchestration into one process. You can run it on a laptop, in a datacenter or in the cloud. It's open source under MIT.
88.8KUpdated 2 months agoMIT
macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API
NextChat is a self-hosted AI chat interface for people who want one place to use their own LLM server and cloud models. The web and desktop project is open source under the MIT license. You can host it with Docker or on Vercel, and desktop clients run on macOS, Windows and Linux.
8KUpdated 11 months agoMIT
Docker#Hybrid search#Knowledge graphs#Multi-user access
R2R is a self-hosted AI retrieval system for developers building applications that answer questions using their own documents. It combines search, retrieval-augmented generation (RAG) and a reasoning agent behind a REST API. The project is open source under the MIT license and runs as a Python service or in Docker.
7.3KUpdated 1 year agoApache-2.0
#Hugging Face integration#Tool calling#Web search
InternLM is a family of downloadable language models for developers and researchers building their own AI applications. It includes models for general conversation, complex reasoning and coding, with separate base and chat variants for customization or use in an assistant.
1.1KUpdated 7 months agoAGPL-3.0
Web · Browser Extension#Multilingual#Multimodal input#Ollama integration
NativeMind brings a local AI assistant into Chrome for people who want help with webpages, documents, and writing without sending that content to a cloud model. The browser extension connects to Ollama on your machine, keeping prompts and AI processing on-device. It requires no account and uses the AGPL-3.0 open-source license.
chatwise.appChat With Your Documents
#MCP#Multimodal input#Tool calling
ChatWise is a desktop chat app for connecting different model providers through one interface. It stores data locally; model processing follows the provider you choose. Its official documentation includes Ollama for models running on your own machine, alongside cloud providers connected with your API keys.
23.8KUpdated 19 hours agoMPL-2.0
macOS · Windows · Linux · iOS · Android#Multilingual#Persistent memory#RAG
Brave Leo is an AI assistant built into the Brave browser, with Bring Your Own Model support for people who want to use their own local or remote models while browsing. It can work with third-party APIs as well as Brave's hosted model choices. The browser runs on macOS, Windows, Linux, Android, and iOS.
3.2KUpdated 5 days agoApache-2.0
macOS · Linux · Docker#GGUF#llama.cpp backend#MCP
Harbor is a CLI and companion app for people experimenting with AI on their own hardware. It manages a local LLM development environment, connecting model backends to chat interfaces and supporting services so you don't have to configure each connection yourself. It's open source under Apache 2.0.
37.5KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#LLM tracing#MCP#Multimodal input
Claude Code Router is an open-source local model gateway for developers who use coding agents and want to manage their model providers in one place. It runs on macOS, Windows and Linux, with Docker and a CLI with a browser interface also available. The project uses the MIT license.
360Updated 3 weeks agoAGPL-3.0
#Ollama integration#RAG#Semantic search
Joplin Jarvis is an AI assistant for people who keep their writing, research, or personal knowledge in Joplin on desktop and mobile. It connects conversations to your existing notes and supports local LLMs through Ollama alongside cloud services such as GPT, Claude, Gemini, and Hugging Face.
2KUpdated 6 months agoAGPL-3.0
macOS · Windows · Linux#Code execution#LM Studio integration#MCP
Witsy is an AGPL-3.0 desktop AI assistant for macOS, Windows and Linux that connects MCP tools to local and cloud models. It's for people who want document chat, writing help and voice features in one app, with a choice of where their models run.
18KUpdated 1 day agoApache-2.0
Web#Code execution#LM Studio integration#MCP
LangBot is a self-hosted AI agent platform for teams that want bots in the messaging apps their customers, coworkers or communities already use. It connects Slack, Discord, Telegram, WeChat and other chat services to models and AI workflows, with a browser dashboard for managing bots across platforms.
enconvo.comAI Workflow Automation
macOS · iOS#LM Studio integration#MCP#MLX
Enconvo is a native AI assistant and agent for Mac that can use your screen and selected text as context, then work inside your apps. It's for people who want help with writing, research and everyday tasks alongside the app they're using. A sidebar keeps the agent beside the current app, while text selection tools give quick access to editing and translation.
17.8KUpdated 2 weeks agoApache-2.0
#Code execution#Hugging Face integration#Human approval
CAMEL is an open-source Python framework for developers and researchers building systems where AI agents work together. Its focus is on agent roles, communication, and behavior across extended tasks, with applications in synthetic training data, task automation, and simulated societies. It uses the Apache 2.0 license.
13.8KUpdated 1 month agoApache-2.0
Browser Extension#Multi-agent workflows#Ollama integration#OpenAI-compatible API
Nanobrowser is an open-source AI agent that automates web tasks inside Chrome or Edge. It's for people who want to delegate repetitive browsing or research while choosing which models handle the work. The extension runs in your browser and is an alternative to OpenAI Operator, with an Apache 2.0 license.