ChuanhuChatGPT is a self-hosted web chat interface for people who want local models and cloud AI services in the same application. It runs on your computer or server and opens in a browser, with support for Windows, macOS and Linux. The Python application is open source under GPL-3.0.
Local model support includes ChatGLM, LLaMA with LoRA models, Qwen, StableLM and MOSS. It can also connect to custom local inference services. Cloud connections include OpenAI's GPT models, Azure OpenAI, Claude, Google Gemini and DeepSeek R1. Hosting the interface yourself doesn't make those API calls local: cloud models process requests through their providers, while local model inference runs on your own hardware.
Beyond chat, ChuanhuChatGPT can answer questions about uploaded files through a knowledge base and use web search to bring online information into responses. Its agent assistant handles tasks automatically in an approach similar to AutoGPT. It also supports GPT-3.5 fine-tuning and image generation through DALL·E 3.
Chats are saved automatically, with separate histories for each user. You can search and rename conversations, or let a model name them. Prompt templates and system prompts support recurring tasks and role-based conversations, and responses can include rendered equations, tables and highlighted code.
The interface adapts to phones and supports installation as a PWA through Chrome, Edge and Safari. A hosted browser demo is available on Hugging Face Spaces.
Claim this page and we'll verify you by hand. ChuanhuChatGPT gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find ChuanhuChatGPT?Promote it
Something wrong or outdated on this page?
12KUpdated 12 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access
h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
32.3KUpdated 1 day ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
37.6KUpdated 2 months agoAGPL-3.0
Web#RAG#Semantic search#Web search
208Updated 4 months ago
Docker · Web#Ollama integration#RAG#Semantic search
SapienAI is a self-hosted academic chatbot and writing workspace for researchers who want their sources, drafts and AI conversations in one place. It runs through Docker and connects to local models through Ollama or cloud models from OpenAI, Anthropic and Google.
11.9KUpdated 2 months agoAGPL-3.0
Web#Code execution#Multilingual#RAG
636Updated 2 months agoMIT
Web#Hugging Face integration#Human approval#Multilingual
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
Khoj is an AI assistant for people who want to ask questions across their own files, research the web, and give recurring work to agents. You can self-host it on your computer or server, or use Khoj's cloud app. It's open source under the GNU AGPL v3.0 license.
Scira is an AI search engine for people who want answers backed by sources and control over the application they use for research. It has a hosted website, and you can self-host the open-source application under AGPL-3.0. Its live search relies on internet access and external services; self-hosting doesn't make that research offline.
ChatLLM Web runs AI chat and a local agent workspace entirely in a WebGPU-capable browser. It's for people who want to work with their own text and code while keeping conversations, selected files, and generated artifacts on their device. The project is open source under the MIT license.