29.4KUpdated 4 days agoAGPL-3.0
iOS · Android · Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Karakeep is a bookmark manager for people who collect links, notes, images and PDFs in one place. You can run the open source app on your own server with Docker under the AGPL-3.0 license, or use the managed Karakeep Cloud service. It has a web app, iOS and Android apps, and extensions for Chrome, Firefox and Safari.
31.3KUpdated 1 day agoApache-2.0
Docker#Hybrid search#Knowledge graphs#llama.cpp backend
Graphiti is a self-hosted Python framework for developers building AI agents that need to remember changing facts. It builds knowledge graphs from conversations, structured records and unstructured text, so an agent can query current information or recover what was true earlier. It's open source under Apache 2.0.
28.4KUpdated 21 hours ago
Web#Human approval#LLM tracing#MCP
Mastra is a TypeScript framework for developers building AI agents and applications on their own servers or inside existing web apps. Its server runs locally or as a standalone deployment; Mastra Cloud provides a hosted alternative. Model routing connects to providers such as OpenAI, Anthropic, and Gemini, so running the framework locally doesn't keep those model requests on your machine.
32.3KUpdated 1 day ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
29.2KUpdated 1 day agoAGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#Multi-user access#Role-based access#Semantic search
Ente Photos is a Google Photos alternative for people who want encrypted photo backups and AI search without giving a storage provider access to their library. You can use its hosted service or self-host the server. Face detection, grouping and natural-language search run on your device; the hosted service stores your encrypted backups.
34.9KUpdated 4 weeks agoApache-2.0
Docker · Web#Agent Skills#Hybrid search#RAG
Qdrant is a self-hosted vector database for developers building semantic search, retrieval-augmented generation (RAG), recommendations and AI agent memory. It stores embeddings alongside JSON metadata, so applications can find similar content while restricting results by attributes such as location, text or numeric ranges. The Rust engine is open source under Apache 2.0 and runs locally in Docker or on your own servers.
39.9KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#Knowledge graphs#LLM tracing#Multimodal input
LightRAG combines knowledge graphs with vector search to answer questions across a document collection. It's a self-hosted Python framework for developers building document assistants, particularly where answers depend on relationships between facts in different files, such as legal or financial material.
91.5KUpdated 7 hours agoApache-2.0
macOS · Windows · Linux · Docker#Hybrid search#MCP#Multi-agent workflows
RAGFlow is an Apache 2.0 licensed RAG engine for teams building AI agents that need to answer questions from their own documents. It can run on a self-hosted server through Docker on Windows, macOS or Linux. A separate hosted cloud service is available.
39.6KUpdated 3 weeks agoMIT
Docker · Web#LM Studio integration#Ollama integration#OpenAI-compatible API
Open Notebook is a self-hosted alternative to Google's NotebookLM for researchers, students, and professionals who want AI help with their own research materials. It runs locally or on your server through Docker and is open source under the MIT license. You choose which content the AI can access and which provider processes it.
115.4KUpdated 5 hours agoAGPL-3.0
iOS · Android · Web#Multi-user access#Semantic search
Immich is a self-hosted photo and video library for people who want their pictures on their own server. Its mobile apps for iOS and Android back up photos and videos, while the web interface gives each user a place to browse and manage their collection. The project is open source under the GNU AGPL v3 license.
66.4KUpdated 5 days agoApache-2.0
Browser Extension#Persistent memory#Semantic search
Mem0 is a memory layer for developers building AI agents and assistants that need to recall earlier interactions. It retains user, session, and agent context across conversations, so an assistant can remember preferences or a support bot can refer to past tickets. The self-hosted code is open source under Apache 2.0; Mem0 also has a managed service.
14Updated 20 hours agoMIT
Windows · Linux · Docker · Web#Human approval#LM Studio integration#Multilingual
ScribeDog is a Markdown editor for writers and note-takers who want AI assistance while keeping their documents on their own hardware. It displays formatted text, tables and images instead of Markdown syntax, but saves ordinary .md files that work with Git and other editors. It has a native desktop app and a self-hosted Server Edition available through Docker.
30.1KUpdated 3 weeks agoMIT
macOS#GGUF#Hugging Face integration#Hybrid search
QMD is a local search engine for people with Markdown notes, meeting transcripts, or documentation they want to search themselves or make available to an AI agent. It accepts exact keywords and natural-language queries, with indexing and model inference running on your own machine. It's open source under MIT.
8.9KUpdated 10 hours agoMIT
macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual
OpenWhispr is a free, MIT-licensed dictation and meeting transcription app for people who want voice input across their apps with control over where processing happens. It's available on macOS, Windows, Linux and iOS. Local transcription works offline and keeps audio on your device; optional cloud transcription sends audio to the selected provider, whose retention policies apply.
recurse.chatChat With Your Documents
macOS#GGUF#Hugging Face integration#OpenAI-compatible API
RecurseChat is a paid Mac app for people who want to chat with AI and ask questions about their files on their own computer. It runs local LLMs without an internet connection or a server, keeping local conversations on the device. The same app also connects to Claude and ChatGPT, so you can choose local processing or a cloud provider.
1.5KUpdated 3 years agoApache-2.0
Web#Guardrails#Semantic search
Rebuff is a prompt injection detector for developers building LLM applications that accept untrusted input. It combines checks for suspicious prompts with a record of past attacks and tests for leaked prompt content. The project is archived and no longer maintained.
496Updated 3 years agoApache-2.0
Docker · Web#Guardrails#Semantic search
Vigil is a self-hosted security scanner for developers and researchers who want to check LLM inputs and responses for prompt injection, jailbreak attempts, and other suspicious content. It combines several detection methods and includes attack signatures and datasets, so teams can assess known threats without building every detector themselves. It is experimental alpha software for research and is open source under Apache 2.0.
24Updated 3 weeks ago
Docker · Web#Hybrid search#Reranking#Scheduled tasks
MindsDB Query Engine is a self-hosted server for developers building AI agents that need access to business data and documents. It queries connected sources through one SQL dialect and adds semantic search for unstructured content. You can run it locally or on your own server with Docker. Its core uses the Elastic License 2.0, with separate licenses in some directories.
5.7KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Ollama integration
OpenAgent is a self-hosted personal AI assistant that combines document search with agents that can act on your behalf. It's for people who want an assistant on their own computer or server, and teams building agents around their documents and workflows. It runs natively on Windows, macOS and Linux as a single executable. It's free and open source under Apache 2.0.
12KUpdated 12 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access
h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
8.5KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux#Ollama integration#OpenAI-compatible API#RAG
Reor is a local AI note-taking app for people who want to search their own writing, find connections between ideas and ask questions about their notes. The project is archived and no longer maintained. It runs on macOS, Windows and Linux, and it's open source under AGPL-3.0.
14.2KUpdated 1 week agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
QAnything is a self-hosted knowledge base for people and teams who want to ask questions about their own documents, including collections that mix Chinese and English. It can answer in either language regardless of the document's language, and runs locally through Docker on Windows, macOS and Linux.
12.2KUpdated 1 month agoMIT
#Multilingual#Multimodal input#Semantic search
FlagEmbedding is an open-source Python toolkit for developers building semantic search or retrieval-augmented generation (RAG) into their own applications. It runs BGE embedding and reranking models, with tools to fine-tune both and evaluate retrieval results. The library uses the MIT license.
7.7KUpdated 4 months agoBSD-3-Clause
Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
Verba is a self-hosted document chatbot for people who want to ask questions across their files and knowledge bases. The project is archived and no longer maintained. It uses Weaviate to find relevant passages and gives those passages to a language model to generate answers.
getzep/zepAgent Memory
Docker#Hybrid search#Knowledge graphs#OpenAI-compatible API
Zep Community Edition v1.0.2 is a deprecated, unsupported self-hosted memory service for developers building AI agents and conversational assistants. This legacy edition uses the Apache 2.0 license. It turns chat history into a knowledge graph that records how facts change over time, so an assistant can distinguish a user's current preferences from earlier ones.
11.9KUpdated 2 months agoAGPL-3.0
Web#Code execution#Multilingual#RAG
Scira is an AI search engine for people who want answers backed by sources and control over the application they use for research. It has a hosted website, and you can self-host the open-source application under AGPL-3.0. Its live search relies on internet access and external services; self-hosting doesn't make that research offline.
2.9KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux · Web#Semantic search#Works offline
OpenRecall records your screen at regular intervals and makes that history searchable with local AI. It's a free, open-source alternative to Microsoft's Windows Recall and Rewind.ai for people who want to find something they previously saw on their computer. It runs on Windows, macOS and Linux, with a browser interface served from your own machine.
3.5KUpdated 4 months agoBSD-3-Clause
VS Code · JetBrains#Code execution#Persistent memory#RAG
Refact is a self-hosted AI coding assistant for developers who want an agent to work across their repository and development tools. The original project is archived and no longer maintained; ongoing development has moved to JegernOUTT/refact. The original code uses the BSD-3-Clause license.
3.7KUpdated 1 day agoMIT
Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend
Twinny is an AI coding assistant for VS Code that lets developers choose where their models run: on their own computer, a private server or a hosted API. It's for individuals and teams who want code suggestions and repository chat with control over where their code goes. The extension and team gateway are open source under the MIT license.
2.8KUpdated 1 month agoMIT
macOS#Batch processing#Hugging Face integration#LoRA
ColPali is a local AI document retrieval library for developers and researchers building document search or retrieval-augmented generation systems. It searches pages as images, using their text, charts and layout together rather than relying on a separate OCR pipeline. The colpali-engine package is deprecated; its maintainers recommend Sentence Transformers for new projects and production use.
6KUpdated 6 months agoMIT
Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
Paperless-AI is a self-hosted extension for Paperless-ngx users who want automatic document sorting and chat with their archive. It requires an existing Paperless-ngx instance and runs in Docker, with a browser interface for reviewing and processing documents. The project is no longer maintained.
5.8KUpdated 4 months agoPostgreSQL
Docker#Batch processing#Ollama integration#RAG
pgai keeps search embeddings in sync with PostgreSQL data for developers building RAG applications and AI agents. It's a Python library with database components and workers you can self-host, including in Docker. The project is archived and no longer maintained or supported. Its code is open source under the PostgreSQL License.
1.1KUpdated 3 weeks agoMPL-2.0
Linux#Ollama integration#RAG#Semantic search
chromem-go is a vector database that runs inside your Go application, so developers can add semantic search or retrieval augmented generation (RAG) without maintaining a separate database server. It stores text alongside embeddings and retrieves related documents for use in LLM answers. Its focus is ordinary application workloads rather than collections containing millions of documents.
39.6KUpdated 1 year ago
#Ollama integration#RAG#Reranking
Quivr Core is a Python framework for developers adding document-based AI answers to their own applications. It combines file ingestion with retrieval-augmented generation (RAG), so a model can answer questions using material from your documents. It supports local models through Ollama as well as cloud APIs from OpenAI, Anthropic, and Mistral.
4.4KUpdated 7 months agoApache-2.0
Docker · Web#Batch processing#Multimodal input#Ollama integration
Cognita is a self-hosted RAG framework for developers building applications that answer questions using their own documents. The project is archived and no longer maintained. It combines a browser interface for document Q&A with reusable components built on LangChain and LlamaIndex, under the Apache 2.0 open-source license.
8.4KUpdated 21 hours agoMIT
#Agent Skills#Batch processing#Guardrails
OGX, formerly Llama Stack, is a self-hosted AI application server for developers building chat apps, document search or AI agents. It brings model inference, file storage, vector search and agent orchestration into one process. You can run it on a laptop, in a datacenter or in the cloud. It's open source under MIT.