RAGapp is a self-hosted app for teams that want an AI assistant using retrieval-augmented generation (RAG) on their own infrastructure. It pairs a browser chat interface with an admin interface for configuring the assistant, taking an approach similar to OpenAI's custom GPTs. It's open source under the Apache 2.0 license.
The app runs in Docker on your own hardware or cloud infrastructure. You can use local models through Ollama or choose hosted models from OpenAI and Gemini. That choice matters for where the AI runs: Ollama provides local model execution, while the OpenAI and Gemini options rely on external model services. Hosting the app yourself doesn't make those cloud models local.
Built on LlamaIndex, RAGapp includes an API alongside its chat interface, so teams can use the assistant directly in a browser or connect it to another application. Its Docker Compose deployments cover an Ollama and Qdrant setup, as well as multiple RAGapp instances with a shared management interface.
Authentication is a separate concern. The standalone container has no built-in authentication layer; it relies on an API gateway to control incoming access. Teams choosing it for an internal service need to provide that access layer as part of their deployment.
Claim this page and we'll verify you by hand. RAGapp gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find RAGapp?Promote it
Something wrong or outdated on this page?
38.7KUpdated 11 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Langchain-Chatchat is a self-hosted application for asking questions about your own documents and using AI agents. It focuses on Chinese-language use and open models, with a fully offline setup that can keep documents and model processing on your hardware. Its code is open source under Apache 2.0.
22.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
32.3KUpdated 24 hours ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
37.5KUpdated 2 months agoAGPL-3.0
Web#RAG#Semantic search#Web search
41.2KUpdated 1 day agoAGPL-3.0
macOS · Docker · Web#Code execution#Hybrid search#LM Studio integration
localGPT is a self-hosted AI document chat app for people who want to question and summarise files on their own hardware. Its local Ollama setup keeps documents and conversations on your machine. Answers include source passages, so you can check what the model used.
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
Khoj is an AI assistant for people who want to ask questions across their own files, research the web, and give recurring work to agents. You can self-host it on your computer or server, or use Khoj's cloud app. It's open source under the GNU AGPL v3.0 license.
AstrBot brings AI assistants into messaging apps such as Telegram, Discord, Slack, QQ and WeCom. It's an open source platform under AGPL-3.0 for people building personal companions, customer support bots or team automation. You can run it on your own computer or server, including through Docker, or use its desktop app for browser-style chat.