LEANN is a local vector database for people building AI search over personal files, research collections or codebases. Its main distinction is a smaller search index: it computes embeddings on demand rather than storing every embedding, reducing the disk space needed for retrieval-augmented generation (RAG).
It runs on macOS, Windows and Linux and is open source under the MIT license. Linux supports CPU-only use. Local model backends can keep document processing and answers on your machine, and LEANN collects no telemetry. It also supports cloud providers such as OpenAI and Anthropic, so privacy depends on which embedding and generation services you choose.
LEANN searches by meaning across PDFs, text and Markdown files, Apple Mail, browser history and conversations from WeChat, iMessage, ChatGPT and Claude. You can ask questions over indexed material or give a coding assistant access to relevant code through its Claude Code MCP integration. MCP connections also bring in live data from Slack and Twitter; exported data can work offline.
For local generation, it works with HuggingFace and Ollama, plus OpenAI-compatible servers such as LM Studio, vLLM and llama.cpp. Embedding support includes sentence-transformers, MLX and Ollama. ColQwen2 and ColPali add PDF retrieval that considers figures, diagrams and page layout alongside text, with MPS acceleration on Apple Silicon. You can transfer indexes between devices.
Claim this page and we'll verify you by hand. LEANN gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find LEANN?Promote it
Something wrong or outdated on this page?
9.1KUpdated 23 hours agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access
Local Deep Research is a self-hosted AI research assistant for people who need cited answers drawn from academic papers, the web and their own documents. It can produce a quick summary or pursue a complex question through repeated searches, then assemble a structured report. It's open source under MIT.
247Updated 2 days agoAGPL-3.0
macOS · Windows · Linux · Docker#Hybrid search#MCP#Multimodal input
2.6KUpdated 16 hours agoApache-2.0
macOS · Windows · Linux#Hybrid search#Knowledge graphs#MCP
XERJ is a local search engine for AI agents that retrieves relevant code and documents instead of making an agent read whole files into its context. It's for developers building coding assistants, codebase Q&A or agents that need persistent memory. The open-source Rust engine runs on Linux, macOS and Windows, on a laptop or self-hosted server, under the Apache 2.0 license.
32.3KUpdated 22 hours ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
3.8KUpdated 7 hours agoApache-2.0
Docker · Web#Knowledge graphs#MCP#Multi-user access
30.1KUpdated 3 weeks agoMIT
macOS#GGUF#Hugging Face integration#Hybrid search
Marginalia is a local-first AI research agent for people whose research papers, notes and working documents are scattered across different formats. It checks relevant passages in the original files before writing answers with citations. Your library stays in readable folders by default.
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
PipesHub is a self-hosted workplace AI platform for teams whose knowledge sits across business apps. It combines company-wide search, cited answers, and agents that can act on connected systems. It's an open source Glean alternative under the Apache 2.0 license, with Docker deployment on your own machine or server.
QMD is a local search engine for people with Markdown notes, meeting transcripts, or documentation they want to search themselves or make available to an AI agent. It accepts exact keywords and natural-language queries, with indexing and model inference running on your own machine. It's open source under MIT.