h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
With local models and its offline document database, chat and document processing can stay on your hardware. It also connects to cloud providers such as OpenAI, Anthropic, Google and Groq; requests sent to those services leave the local setup. Supported local backends include Ollama, llama.cpp, GPT4ALL and vLLM, with models such as Mistral, Mixtral and LLaMa2. It can use CPUs, CUDA GPUs and Apple Silicon.
Document Q&A covers PDFs, Word files, spreadsheets, code and Markdown, plus images, audio and video frames. You can keep personal or shared document collections, search their contents, and generate summaries or extract information across documents. OCR support helps it read text from scanned material.
The browser interface includes streaming responses and side-by-side model comparisons. Beyond text chat, h2oGPT supports image understanding with LLaVa, image generation with Stable Diffusion and Flux, and Whisper speech transcription. Spoken responses, voice cloning and hands-free voice control are also available.
For developers, an OpenAI-compatible API exposes chat, embeddings, audio and image generation to other applications. Open WebUI can use h2oGPT as its backend. API-based agents handle web research, document questions and Python code execution, including generating plots.
Claim this page and we'll verify you by hand. h2oGPT gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find h2oGPT?Promote it
Something wrong or outdated on this page?
7.1KUpdated 23 hours agoMIT
Docker · Web#LM Studio integration#Multimodal input#Ollama integration
big-AGI is a browser-based AI workspace for researchers, developers and people who want to compare model answers before relying on them. Its Beam feature sends the same prompt to several models in parallel, without letting them see each other's replies. You can examine disagreements, then use Merge to combine the responses into a single answer with a preset or custom prompt.
4.8KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
47.7KUpdated 1 month agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#GGUF#llama.cpp backend#LoRA
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
1.6KUpdated 1 year agoMIT
Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input
41.2KUpdated 1 day agoAGPL-3.0
macOS · Docker · Web#Code execution#Hybrid search#LM Studio integration
Lollms WebUI is a local, single-user AI interface for people who want text chat and media generation in one place. It runs on Windows, macOS and Linux, with Docker support, and lets writers, developers and other users choose models and task-specific personalities. It's free and open source under Apache 2.0. The project receives minimal maintenance.
text-generation-webui, also called TextGen, runs language models on your own hardware through a desktop app or a self-hosted browser interface. It's for people who want private chat and writing tools, and developers who need a local model API. It works offline without telemetry; web search and page fetching use the internet.
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
Amica is a locally runnable interface for talking with customizable 3D AI characters. It's for people who want an animated, voiced character as the face of their AI assistant, with a choice of local LLM backends or cloud services. The project builds on Pixiv's ChatVRM.
AstrBot brings AI assistants into messaging apps such as Telegram, Discord, Slack, QQ and WeCom. It's an open source platform under AGPL-3.0 for people building personal companions, customer support bots or team automation. You can run it on your own computer or server, including through Docker, or use its desktop app for browser-style chat.