544Updated 14 hours agoMIT
macOS · iOS · Web#Code execution#Distributed execution#Hugging Face integration
Pooled runs a single open model across browser tabs on laptops, desktops and phones, combining their memory when the model won't fit on one device. It's for people who want local AI chat or a coding assistant using hardware they already have. It's open source under the MIT license and requires no account or per-device installation.
49.8KUpdated 23 hours agoMIT
macOS · Windows · Web#Ollama integration
AIRI is a self-hosted AI companion for people who want a virtual character they can talk to and play games with. Inspired by Neuro-sama, it combines real-time voice conversation with game integrations for Minecraft and Factorio. The project is in early development; Factorio support is a work in progress with a proof of concept. It's open source under the MIT license.
390.8KUpdated 19 hours ago
macOS · Windows · Linux · iOS · Android · Docker#Multi-user access#Persistent memory#Tool calling
OpenClaw is a self-hosted AI assistant that connects your own computer to the messaging services you already use. It's for people who want a personal assistant on their laptop, or teams that want to run a shared assistant on their own hardware. The project is open source under the MIT license.
32.3KUpdated 24 hours ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
45.2KUpdated 5 hours agoMIT
Web#Agent Skills#Code execution#MCP
LibreChat is a self-hosted ChatGPT alternative for people and teams that want local models and cloud AI services in one chat interface. It runs on a server you control and is open source under the MIT license. You can switch models without moving to a separate chat app.
130KUpdated 40 minutes agoMIT
Web#Code execution#GGUF#Hugging Face integration
llama.cpp runs language models on your own hardware and can serve them from a machine you control. It’s an MIT-licensed, open source inference engine for people building local AI apps, running a private model server, or using a model directly from the command line. It supports vision-language models too.
153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
23.8KUpdated 8 months agoMIT
Web#LLM tracing#Multi-user access#Ollama integration
Vanna is a self-hosted Python framework for building AI agents that answer database questions in plain language. It's for teams adding chat to analytics products or internal data tools, especially when each user needs different access to the data. The project is archived and no longer maintained. Its open-source code uses the MIT license.
12KUpdated 12 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access
h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
1.1KUpdated 5 months agoMIT
Web · Browser Extension#Ollama integration
ollama-ui is a small browser interface for people who run models through Ollama and want a graphical place to chat. You can serve the interface locally and open it in a browser, or use the Chrome extension. It's open source under the MIT license.
4.4KUpdated 2 years agoApache-2.0
Docker · Web#Ollama integration#RAG
RAGapp is a self-hosted app for teams that want an AI assistant using retrieval-augmented generation (RAG) on their own infrastructure. It pairs a browser chat interface with an admin interface for configuring the assistant, taking an approach similar to OpenAI's custom GPTs. It's open source under the Apache 2.0 license.
11.9KUpdated 2 months agoAGPL-3.0
Web#Code execution#Multilingual#RAG
Scira is an AI search engine for people who want answers backed by sources and control over the application they use for research. It has a hosted website, and you can self-host the open-source application under AGPL-3.0. Its live search relies on internet access and external services; self-hosting doesn't make that research offline.
35Updated 7 months agoMIT
macOS · Web · VS Code#Code execution#LM Studio integration#MCP
Wingman-AI is a self-hosted AI agent platform for work that needs ongoing context and several agents with different roles. It suits developers and teams handling research, support, or recurring operations alongside coding tasks. A lead agent can delegate work to specialized subagents, each with its own workspace and session history.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
14KUpdated 5 months ago
macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP
Open-LLM-VTuber is a local AI companion for people who want a character they can talk to, with a Live2D avatar that responds through speech and expressions. It runs on Windows, macOS and Linux through web and desktop clients. With local models for speech and language processing, it works fully offline and keeps conversations on your device. Cloud APIs are optional alternatives that send the corresponding processing to external services.
1.6KUpdated 1 year agoMIT
Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input
Amica is a locally runnable interface for talking with customizable 3D AI characters. It's for people who want an animated, voiced character as the face of their AI assistant, with a choice of local LLM backends or cloud services. The project builds on Pixiv's ChatVRM.
752Updated 9 months agoAGPL-3.0
Web#llama.cpp backend#OpenAI-compatible API#Persistent memory
Mikupad is a browser-based LLM frontend packaged as a single HTML file, for writers and people who want control over generated text. It connects to a model backend of your choice and supports both direct text continuation and chat with instruct models. It's open source under AGPL-3.0.
784Updated 4 months agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multi-user access#Multimodal input#Persistent memory
Agnai is a self-hosted AI roleplay chat app for people who want to create fictional characters and talk with them alone or in a group. A conversation can include multiple people and multiple bots. It builds on early work from Galatea-UI by PygmalionAI and uses the AGPL-3.0 open-source license.
1.7KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#Multilingual#Persistent memory
RisuAI is an open-source AI roleplay client for people who want to create characters, build fictional worlds and chat with several characters together. It runs on Windows, macOS, Linux, Android and in a browser. You can also host the web app yourself with Docker.
5.7KUpdated 1 year agoApache-2.0
Windows · Docker · Web#llama.cpp backend
Serge is a self-hosted chat interface for people who want to run language models on their own hardware and talk to them in a browser. The project is archived and no longer maintained. It uses llama.cpp to run models locally, with Alpaca as a named chat model and LLaMA also referenced in its memory requirements.
10.9KUpdated 3 years agoMIT
macOS · Docker · Web#GGUF#llama.cpp backend#OpenAI-compatible API
LlamaGPT is a self-hosted ChatGPT alternative for people who want general chat or coding help on their own computer or home server. It runs models locally and keeps conversation data on your device. After the initial model download, it works offline.
4.8KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
Lollms WebUI is a local, single-user AI interface for people who want text chat and media generation in one place. It runs on Windows, macOS and Linux, with Docker support, and lets writers, developers and other users choose models and task-specific personalities. It's free and open source under Apache 2.0. The project receives minimal maintenance.
33.3KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Ollama integration
Chatbot UI is a browser-based AI chat app for people who want to host their own interface and choose between local models and cloud providers. It connects to Ollama for models running on your hardware, alongside services such as OpenAI, Azure OpenAI, Anthropic and Google Gemini. The app is open source under the MIT license.
88.8KUpdated 2 months agoMIT
macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API
NextChat is a self-hosted AI chat interface for people who want one place to use their own LLM server and cloud models. The web and desktop project is open source under the MIT license. You can host it with Docker or on Vercel, and desktop clients run on macOS, Windows and Linux.
25KUpdated 2 years agoApache-2.0
macOS · Web#LoRA#Multimodal input#Quantization
LLaVA is a family of vision-language models for researchers and developers who want to ask questions about images on their own hardware. It pairs a CLIP vision encoder with a language model to support image descriptions, visual reasoning and reading text in pictures. Its Python code is open source under Apache 2.0; the project places research-use restrictions on its data and checkpoints, with additional terms from the underlying models.
1.2KUpdated 12 months agoMIT
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Hollama is an open-source LLM chat app whose interface runs entirely in your browser. It's for people who want to chat with local AI through Ollama or connect to OpenAI servers, with support for multiple server connections. The interface stores data locally in the browser; the connected server handles model requests, so where inference runs depends on the server you choose.
8.2KUpdated 3 days agoMIT
Web · Browser Extension#LM Studio integration#Ollama integration#OpenAI-compatible API
Page Assist brings local AI chat into your browser, with a sidebar for conversations alongside a webpage and a separate tab for a ChatGPT-style interface. It's for people who already run models on their own hardware and want to ask questions about what they're reading without leaving the page.
11KUpdated 1 day agoApache-2.0
Docker · Web#llama.cpp backend#MCP#Multi-user access
HuggingChat UI is the open-source chat application behind Hugging Face's hosted HuggingChat. You can run it on your own computer or server and connect it to a local LLM backend or a cloud provider. It's for people and teams who want a browser-based ChatGPT alternative with control over the chat service and where its data lives. The code uses the Apache 2.0 license.
1.5KUpdated 3 weeks agoMIT
Docker · Web#Ollama integration#OpenAI-compatible API#Streaming inference
Unmute adds spoken conversation to text LLMs using Kyutai's speech recognition and speech synthesis models. It's for developers who want a self-hosted voice interface while keeping their choice of language model. The project uses the MIT license, and a hosted browser demo is available at Unmute.sh.
12.5KUpdated 4 months agoApache-2.0
Docker · Web#Hugging Face integration#OpenAI-compatible API
OpenLLM is a self-hosted LLM server for developers who want to connect their applications to models running on their own hardware or servers. Its OpenAI-compatible API works with clients built for that interface, including the OpenAI Python client and LlamaIndex. The project is open source under the Apache License 2.0.
38.7KUpdated 11 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Langchain-Chatchat is a self-hosted application for asking questions about your own documents and using AI agents. It focuses on Chinese-language use and open models, with a fully offline setup that can keep documents and model processing on your hardware. Its code is open source under Apache 2.0.
41.2KUpdated 1 day agoAGPL-3.0
macOS · Docker · Web#Code execution#Hybrid search#LM Studio integration
AstrBot brings AI assistants into messaging apps such as Telegram, Discord, Slack, QQ and WeCom. It's an open source platform under AGPL-3.0 for people building personal companions, customer support bots or team automation. You can run it on your own computer or server, including through Docker, or use its desktop app for browser-style chat.
3.1KUpdated 2 months agoGPL-3.0
Web#MCP#Multi-agent workflows#Ollama integration
Cheshire Cat is a self-hosted Python framework for people learning how AI agents work or building custom assistants for research and creative projects. It pairs a local web chat interface with an API, so you can test an agent in conversation and use it inside another application. It's open source under GPL-3.0.
22.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
localGPT is a self-hosted AI document chat app for people who want to question and summarise files on their own hardware. Its local Ollama setup keeps documents and conversations on your machine. Answers include source passages, so you can check what the model used.
47.7KUpdated 1 month agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#GGUF#llama.cpp backend#LoRA
text-generation-webui, also called TextGen, runs language models on your own hardware through a desktop app or a self-hosted browser interface. It's for people who want private chat and writing tools, and developers who need a local model API. It works offline without telemetry; web search and page fetching use the internet.
25.8KUpdated 4 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access
kotaemon is a self-hosted document chat app for people who want to ask questions across their files and check where the answers came from. It runs in a browser on Windows, macOS or Linux, with Docker also supported. The project uses the Apache 2.0 license.