153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
24.9KUpdated 1 week agoMIT
Windows · Docker · Web#Batch processing#ONNX
rembg removes image backgrounds on your own hardware, with batch processing and a Python library for developers building image workflows. It's also useful for people preparing product cutouts or portraits who want control over processing and model choice. Local models keep images on your machine and can work offline once downloaded.
46.2KUpdated 23 hours agoGPL-3.0
Docker · Web#Role-based access#Single sign-on
Paperless-ngx is a self-hosted document management system for people who want to keep scanned paperwork in a searchable digital archive. It brings document scanning, indexing and storage together, so you can find records by their contents rather than sift through paper folders. You can run it on a local server at home, and it supports deployment with Docker.
49.3KUpdated 2 hours agoMIT
macOS · Linux · Docker · Web#Code execution#Human approval#llama.cpp backend
LocalAI runs language models, speech, vision and image generation on hardware you control. It's for developers and teams that want a self-hosted AI server for their apps without sending model requests to a cloud service. Its OpenAI-compatible API works with existing clients, and it also accepts Anthropic, Ollama and ElevenLabs API calls.
77KUpdated 22 hours agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image
Unsloth brings model training and everyday AI use into a desktop app for people who want to run models on their own hardware. Its no-code interface covers chat, fine-tuning and media generation on macOS, Windows and Linux. The Unsloth software is open source under Apache 2.0.
15KUpdated 11 months ago
macOS · Windows · Linux · Web#Hugging Face integration
Hunyuan3D generates textured 3D assets from reference images or text and can run on your own computer. It's for 3D artists, hobbyists and developers who want AI-generated meshes they can use in other software. It supports macOS, Windows and Linux; Hunyuan3D Studio is a separate hosted option.
picapport.deAI Photo Libraries
macOS · Windows · Linux · Docker · Web#Batch processing#Multi-user access#Role-based access
PicApport is a self-hosted photo server for families and businesses that want a searchable media archive on their own hardware. Its AI tagging add-on recognizes image content without sending photos to a cloud service and saves the resulting tags in image metadata. The server runs on Windows, Linux or macOS, with a Docker image available, and you access the gallery through a browser.
14Updated 20 hours agoMIT
Windows · Linux · Docker · Web#Human approval#LM Studio integration#Multilingual
ScribeDog is a Markdown editor for writers and note-takers who want AI assistance while keeping their documents on their own hardware. It displays formatted text, tables and images instead of Markdown syntax, but saves ordinary .md files that work with Git and other editors. It has a native desktop app and a self-hosted Server Edition available through Docker.
6.5KUpdated 6 days agoAGPL-3.0
Linux · Docker · Web#Multi-user access
Photoview is a self-hosted photo gallery for photographers and households who keep their pictures on a personal server or NAS. It turns existing folders into browser-based albums and automatically scans for added photos and videos. Your folder structure determines how the library appears, so you can keep organizing files through Samba, FTP or Nextcloud. Your media stays on your server.
12KUpdated 1 week agoApache-2.0
Docker · Web#Batch processing#Human approval#Multi-agent workflows
Bisheng is an open source, self-hosted platform for teams building AI applications around business documents and processes. Its visual workflow editor combines automated tasks with human feedback, including intervention during multi-turn conversations. It's suited to document review, support ticket assistance and report generation that need more control than a single chatbot exchange.
prodi.gyData Labeling and Annotation
Web#Hugging Face integration#Works offline
Prodigy is a proprietary annotation tool that runs on your own machines, including air-gapped systems without an internet connection. It's for developers and research teams building training and evaluation datasets for custom AI models. The Python library includes a web application where annotators can label data without programming knowledge.
3.3KUpdated 1 month agoApache-2.0
Docker · Web#Multi-user access#Prompt versioning
Pezzo is an open-source platform for developers and teams managing prompts and monitoring LLM applications. You can run the full stack locally with Docker Compose, keeping the prompt management and monitoring platform on infrastructure you control. Its source code uses the Apache 2.0 license.
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.
4.6KUpdated 19 hours agoMIT
Web#AI red teaming#OpenAI-compatible API
PyRIT is an MIT-licensed, open source Python framework for security professionals and engineers assessing generative AI systems. It combines automated attack testing with human-led investigations through CoPyRIT, a web interface served locally. The framework runs locally, but prompts go to the target services you choose; cloud targets and cloud-based scorers process requests outside your machine.
1.5KUpdated 3 years agoApache-2.0
Web#Guardrails#Semantic search
Rebuff is a prompt injection detector for developers building LLM applications that accept untrusted input. It combines checks for suspicious prompts with a record of past attacks and tests for leaked prompt content. The project is archived and no longer maintained.
496Updated 3 years agoApache-2.0
Docker · Web#Guardrails#Semantic search
Vigil is a self-hosted security scanner for developers and researchers who want to check LLM inputs and responses for prompt injection, jailbreak attempts, and other suspicious content. It combines several detection methods and includes attack signatures and datasets, so teams can assess known threats without building every detector themselves. It is experimental alpha software for research and is open source under Apache 2.0.
1.2KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing
GoModel is a self-hosted AI gateway for developers and platform teams that want one API for local models and cloud providers. It accepts OpenAI- and Anthropic-compatible requests, so applications can keep their existing SDKs while the gateway handles provider selection and usage controls.
37KUpdated 2 years agoMIT
Linux · Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API
One API is a self-hosted LLM API gateway for developers and teams that want to share model access across apps or users. It puts cloud providers and Ollama behind an OpenAI-compatible API, so clients can use one endpoint across different backends. It's open source under MIT and runs on your own server as a single executable or in Docker.
1.4KUpdated 3 weeks ago
Web#Home Assistant integration#OpenAI-compatible API#Tool calling
Extended OpenAI Conversation is a custom component for Home Assistant users who want an AI assistant that can act on their home. It builds on OpenAI Conversation and adds device control, automation creation, and access to historical device states. It runs inside Home Assistant, while a separate model backend handles the conversation.
318Updated 3 weeks agoApache-2.0
macOS · Windows · Linux · iOS · Android · Web#Quantization
picoLLM is an on-device inference SDK for developers building apps that run compressed language models on users' hardware. It generates text locally, so prompts don't need to go to a cloud inference service. Its main distinction is Picovoice's compression method, which learns how to allocate precision across model weights rather than applying a fixed allocation.
3.9KUpdated 1 year agoGPL-3.0
macOS · Windows · Linux · Web#Hugging Face integration#Streaming inference#Voice conversion
Seed-VC changes recorded speech or singing to sound like a voice supplied in a short reference clip, without training a separate model for that speaker. It runs locally on Windows, Linux and Apple Silicon Macs, with uses in audio production, live streaming and online meetings. The project is archived and no longer maintained.
2.9KUpdated 1 year agoAGPL-3.0
Web · Browser Extension
Amurex is an AI meeting copilot that runs as a Chrome extension for Google Meet and MS Teams. It's for people who want help following a discussion and recording decisions without handling all the meeting notes themselves. You can self-host its backend and web application alongside the browser extension.
24Updated 3 weeks ago
Docker · Web#Hybrid search#Reranking#Scheduled tasks
MindsDB Query Engine is a self-hosted server for developers building AI agents that need access to business data and documents. It queries connected sources through one SQL dialect and adds semantic search for unstructured content. You can run it locally or on your own server with Docker. Its core uses the Elastic License 2.0, with separate licenses in some directories.
23.8KUpdated 8 months agoMIT
Web#LLM tracing#Multi-user access#Ollama integration
Vanna is a self-hosted Python framework for building AI agents that answer database questions in plain language. It's for teams adding chat to analytics products or internal data tools, especially when each user needs different access to the data. The project is archived and no longer maintained. Its open-source code uses the MIT license.
5.7KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Ollama integration
OpenAgent is a self-hosted personal AI assistant that combines document search with agents that can act on your behalf. It's for people who want an assistant on their own computer or server, and teams building agents around their documents and workflows. It runs natively on Windows, macOS and Linux as a single executable. It's free and open source under Apache 2.0.
1.4KUpdated 1 year agoMIT
Docker · Web#Home Assistant integration
Double Take connects Frigate camera images to facial recognition services through a shared web interface and API. It's for people running camera systems and Home Assistant who want to identify familiar faces and use the results in home automations. The app runs on your own server in Docker and is open source under the MIT license.
53.2KUpdated 1 year agoGPL-3.0
macOS · Windows · Linux · Docker · Web#ControlNet#Image-to-image#Inpainting
Fooocus is a free, open-source AI image generator for people who want to create images on their own computer without spending much time tuning settings. It uses Stable Diffusion XL and automatically expands prompts with a local GPT-2 engine. Generation works offline once the required models are downloaded, so prompts and images can stay on your machine.
12KUpdated 12 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access
h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
14.2KUpdated 1 week agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
QAnything is a self-hosted knowledge base for people and teams who want to ask questions about their own documents, including collections that mix Chinese and English. It can answer in either language regardless of the document's language, and runs locally through Docker on Windows, macOS and Linux.
3.3KUpdated 2 months agoMIT
Windows · Linux · Docker · Web#Hugging Face integration#LoRA
FluxGym is a local web interface for training FLUX LoRAs on your own images, with support for GPUs with 12GB, 16GB or 20GB of VRAM. It's for people who want to customize an image model through a browser while retaining access to detailed training controls. It runs on Windows and Linux, with Docker support, and is open source under the MIT license.
1.1KUpdated 5 months agoMIT
Web · Browser Extension#Ollama integration
ollama-ui is a small browser interface for people who run models through Ollama and want a graphical place to chat. You can serve the interface locally and open it in a browser, or use the Chrome extension. It's open source under the MIT license.
3.9KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#Multimodal input
Riffusion is a Python library for generating music and audio on your own hardware using Stable Diffusion. It's for developers and musicians who want to experiment with text-driven sound generation or build it into an app. The hobby project is no longer actively maintained.
6.4KUpdated 2 years agoApache-2.0
Web · Browser Extension#Code execution
LaVague is a Python framework for developers building AI agents that carry out tasks in a web browser. An agent takes a plain-language objective, examines the current page, and generates and executes browser actions across multiple steps. It's open source under the Apache 2.0 license.
7.7KUpdated 4 months agoBSD-3-Clause
Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
Verba is a self-hosted document chatbot for people who want to ask questions across their files and knowledge bases. The project is archived and no longer maintained. It uses Weaviate to find relevant passages and gives those passages to a language model to generate answers.
4.4KUpdated 2 years agoApache-2.0
Docker · Web#Ollama integration#RAG
RAGapp is a self-hosted app for teams that want an AI assistant using retrieval-augmented generation (RAG) on their own infrastructure. It pairs a browser chat interface with an admin interface for configuring the assistant, taking an approach similar to OpenAI's custom GPTs. It's open source under the Apache 2.0 license.
6.9KUpdated 1 year agoApache-2.0
Web#Multi-agent workflows#Multilingual#RAG
MindSearch is a self-hosted AI search framework for people who want to build their own Perplexity-style answer engine. Multiple LLM agents search and read web pages to produce answers with a visible research process. It's open source under Apache 2.0.