Self-Hosted AI Tools You Can Run in Docker

AI apps and services that ship a Docker image, so you can start them on a home server, NAS or VPS without installing their dependencies.

200+ tools
Self-hosted AI document search and agents with source citations. Run local models through Ollama, vLLM or llama.cpp, including fully offline deployments.

18.3KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend

An open-source PDF translator that preserves formulas and layouts, runs locally or in Docker, and supports Ollama, Google, DeepL and OpenAI.

37.3KUpdated 1 day agoAGPL-3.0

macOS · Windows · Docker · Web#Batch processing#Hugging Face integration#MCP

Open-source PostgreSQL extension for vector search with pgvector, DiskANN indexing and compression. Run it self-hosted or in Timescale Cloud.

3.1KUpdated 3 weeks agoPostgreSQL

macOS · Linux · Docker#Semantic search

A self-hosted text embedding server with a REST API, CPU and GPU support, and offline operation with downloaded model weights. Apache 2.0 licensed.

5.1KUpdated 1 week agoApache-2.0

macOS · Linux · Docker#Batch processing#Hugging Face integration#LLM tracing

An open-source local AI API under Apache 2.0 that connects to Ollama, llama.cpp and other OpenAI-compatible servers for document retrieval and agent workflows.

57.6KUpdated 1 week agoApache-2.0

Docker · Web#Code execution#llama.cpp backend#MCP

Self-hosted computer vision server for images and video, with Docker support, NVIDIA GPU acceleration, and optional Roboflow hosted compute.

2.5KUpdated 1 day ago

macOS · Windows · Linux · Docker#Batch processing#Code execution#Multimodal input

Local audiobook converter turns ebooks into narrated audio with chapters and voice cloning. Runs on Windows, macOS and Linux under Apache 2.0.

20.3KUpdated 4 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning

Favicon of Lemonade

Lemonade

1 video
An open source local AI server for chat, image generation, and speech on Windows, macOS, and Linux, with APIs for apps and agents.

5.8KUpdated 3 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Docker#GGUF#Hugging Face integration#llama.cpp backend

A self-hosted subtitle translator that runs in Docker and automates media library translation using local Ollama models or cloud services such as DeepL.

881Updated 3 days ago

Docker#Multilingual#Ollama integration

Self-hosted workflow platform turns Python, TypeScript and other scripts into APIs and internal apps. Runs on Docker or Kubernetes; cloud hosting is available.

18.1KUpdated 24 hours ago

Linux · Docker · Web#Code execution#Git integration#Human approval

Favicon of Presidio

Presidio

1 video
A self-hosted Python framework that detects and anonymizes sensitive data in text, images and structured records. Open source under the MIT license.

11.1KUpdated 2 days agoMIT

Docker#Human approval#Multilingual

Self-hosted AI research assistant for web, papers and private documents. Runs on Windows, macOS and Linux with Ollama or cloud models. MIT licensed.

9.1KUpdated 23 hours agoMIT

macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access

Open-source AI framework for semantic search, RAG and agents. Runs locally or in Docker, with Hugging Face, llama.cpp and cloud models via LiteLLM.

13KUpdated 23 hours agoApache-2.0

Docker#Agent Skills#Hugging Face integration#Knowledge graphs

Favicon of Chatwoot

Chatwoot

1 video
Self-hosted customer support platform with an AI agent, shared inbox and help center. Run it on your own server or use the hosted cloud service.

37.3KUpdated 21 hours ago

Docker · Web#Multi-user access#Multilingual#RAG

Open source data labeling and AI evaluation platform that runs locally or on your server, with custom annotation interfaces and model-assisted labeling.

28.4KUpdated 1 day agoApache-2.0

macOS · Windows · Docker · Web#Human approval#Multi-user access#Multimodal input

A self-hosted AI chatbot platform with a visual builder, local LLM support and human handoff. Runs on Docker or Kubernetes, with a hosted option.

322Updated 2 weeks agoMIT

iOS · Android · Docker · Web#Batch processing#Code execution#Guardrails

A self-hosted LLM inference library built on PyTorch for NVIDIA GPUs, with a Python API, OpenAI-compatible serving, and multi-node support.

14.7KUpdated 22 hours ago

Docker#Batch processing#Distributed execution#LoRA

Self-hosted web extraction API converts pages and PDFs to Markdown for LLMs. Apache 2.0 service code runs in Docker; a hosted API is also available.

12.1KUpdated 4 months agoApache-2.0

Docker#Multimodal input#Structured output

Self-hosted speech API for transcription, translation and speech generation. Runs via Docker on CPU or GPU with faster-whisper, Kokoro and Piper.

3.7KUpdated 5 months agoMIT

Docker#OpenAI-compatible API#Streaming inference

An open-source browser automation server for AI agents that runs locally on macOS, Windows or Linux and reads page structure without a vision model.

37.7KUpdated 2 days agoApache-2.0

macOS · Windows · Linux · Docker#Code execution#LM Studio integration#MCP

Self-hosted AI model serving platform for Linux, Windows and macOS. Run language, speech and image models through an OpenAI-compatible API under Apache 2.0.

9.6KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input

Favicon of llama-swap

llama-swap

1 video
A local AI proxy that switches models on demand through OpenAI and Anthropic compatible APIs. Runs on macOS, Windows, Linux and FreeBSD under MIT.

5.8KUpdated 2 days agoMIT

macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend

An open-source local LLM framework that splits work across CPUs and GPUs, with SGLang serving and LlamaFactory fine-tuning under Apache 2.0.

19.5KUpdated 1 week agoApache-2.0

Docker#LoRA#Multimodal input#Prompt caching

A Python library for asking questions about SQL, CSV and parquet data, with chart generation and a separate hosted business intelligence app.

23.8KUpdated 11 months ago

Docker#Code execution#RAG

LLM training framework with ready-made research scripts, NVIDIA GPU parallelism, and Hugging Face checkpoint conversion through Megatron Bridge.

18KUpdated 1 day ago

Docker#Distributed execution#Hugging Face integration#Quantization

Favicon of GPT4All

GPT4All

1 video
An open-source local AI chatbot for Windows, macOS and Linux. Run models without a GPU or cloud API, and chat privately with your documents.

77.4KUpdated 1 year agoMIT

macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#OpenAI-compatible API

Self-hosted subtitle generator runs Whisper locally on CPU or NVIDIA GPU and connects to Bazarr, Plex, Jellyfin, Emby and Tautulli. MIT licensed.

1.5KUpdated 2 months agoMIT

Docker#Batch processing#Multilingual#OpenAI-compatible API

A self-hosted AI agent builder with visual workflows and document retrieval. Runs through Docker, with hosted and commercial editions also available.

29.8KUpdated 1 day ago

Docker · Web#LLM tracing#MCP#Multi-user access

Favicon of Weaviate

Weaviate

1 video
A self-hosted vector database that combines semantic and keyword search with RAG. Run it locally with Docker, on Kubernetes, or in a managed cloud service.

16.9KUpdated 1 day ago

Docker#Hybrid search#RAG#Reranking

A local NVIDIA Jetson monitoring tool with a terminal interface, Python API and Docker support. Open source under AGPL-3.0.

2.6KUpdated 1 week agoAGPL-3.0

Linux · Docker

Favicon of Cognee

Cognee

2 videos
An open-source AI agent memory platform that runs locally on CPU, connects to Claude Code and Codex, and supports self-hosting or managed cloud hosting.

31.2KUpdated 23 hours agoApache-2.0

Docker · Web#Knowledge graphs#MCP#Multi-user access

An MIT licensed tool that packages local or remote repositories for Claude, ChatGPT, Gemini, and MCP assistants through a command line or website.

28.6KUpdated 2 days agoMIT

Docker · Web · Browser Extension#Agent Skills#Code execution#MCP

A self-hosted AI workspace with parallel model chats and answer merging. Connect Ollama, LM Studio or cloud providers using your own API keys.

7.1KUpdated 1 day agoMIT

Docker · Web#LM Studio integration#Multimodal input#Ollama integration

An open-source OCR toolkit that converts PDFs and images into Markdown using a local GPU or an OpenAI-compatible inference server. Apache 2.0 licensed.

19.7KUpdated 6 months agoApache-2.0

Linux · Docker · Web#Batch processing#Distributed execution#OpenAI-compatible API

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

Open-source meeting bot API for Meet, Teams and Zoom, with live transcripts and AI agent access. Self-host with Docker or use the hosted service.

2.8KUpdated 2 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#Git integration#Human approval