localGPT is a self-hosted AI document chat app for people who want to question and summarise files on their own hardware. Its local Ollama setup keeps documents and conversations on your machine. Answers include source passages, so you can check what the model used.
It handles PDF, DOCX, HTML, Markdown and plain text through Docling, with OCR for PDFs that lack a text layer. The browser interface lets you organise document collections and keep conversations by topic. It also shows progress while the app searches your files and generates an answer.
Search combines meaning-based retrieval with LanceDB full-text search, then reranks the matches for relevance. This lets it look for both related ideas and exact wording. It can break complex questions into smaller questions and choose between consulting your documents or answering directly with the model. Optional sentence pruning removes irrelevant material from retrieved passages.
Ollama provides the generation models, with Qwen models used by default. You can choose different models for individual conversations or messages. Embeddings can come from HuggingFace models such as harrier-oss-v1 and Qwen3-Embedding, or from Ollama.
The MIT-licensed project runs on macOS, Windows and Linux, with Docker deployment available. It requires at least 8GB of RAM and recommends 16GB or more. Models need an initial download; embedding and reranking can use NVIDIA CUDA, Apple MPS or the CPU. Developers can use its REST APIs to add document indexing and chat to their own applications.
Claim this page and we'll verify you by hand. localGPT gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find localGPT?Promote it
Something wrong or outdated on this page?
38.7KUpdated 11 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Langchain-Chatchat is a self-hosted application for asking questions about your own documents and using AI agents. It focuses on Chinese-language use and open models, with a fully offline setup that can keep documents and model processing on your hardware. Its code is open source under Apache 2.0.
32.3KUpdated 24 hours ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
37.5KUpdated 2 months agoAGPL-3.0
Web#RAG#Semantic search#Web search
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
41.2KUpdated 1 day agoAGPL-3.0
macOS · Docker · Web#Code execution#Hybrid search#LM Studio integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
Khoj is an AI assistant for people who want to ask questions across their own files, research the web, and give recurring work to agents. You can self-host it on your computer or server, or use Khoj's cloud app. It's open source under the GNU AGPL v3.0 license.
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
AstrBot brings AI assistants into messaging apps such as Telegram, Discord, Slack, QQ and WeCom. It's an open source platform under AGPL-3.0 for people building personal companions, customer support bots or team automation. You can run it on your own computer or server, including through Docker, or use its desktop app for browser-style chat.