Self-Hosted AI Tools You Can Run in Docker

AI apps and services that ship a Docker image, so you can start them on a home server, NAS or VPS without installing their dependencies.

200+ tools
A self-hostable model-sharing platform with accounts, uploads and comments. Its public website adds hosted generation that local setup does not include.

7.3KUpdated 17 hours agoApache-2.0

Linux · Docker · Web#Multi-user access

A coding assistant CLI that writes and runs code from plain-language requests. Runs locally or in Docker, with local models or OpenAI and Anthropic APIs.

55.1KUpdated 2 years agoMIT

Windows · Docker#Code execution#Multimodal input

A self-hosted AI coding agent that plans tasks, researches the web and writes code, with Ollama support and an MIT license.

19.6KUpdated 1 year agoMIT

macOS · Windows · Linux · Docker · Web#Ollama integration#Web search

An open-source image-to-3D tool that runs locally with CUDA, exports OBJ meshes, and includes a Gradio interface and Docker support. Apache 2.0 licensed.

4.5KUpdated 2 years agoApache-2.0

Docker · Web#Image-to-image

A local text embedding model for semantic search and RAG, with adjustable vector sizes, Apache 2.0 licensing, and support for Sentence Transformers.

1.9KUpdated 11 months ago

Docker#Batch processing#Hugging Face integration#ONNX

Self-hosted ML experiment tracker for comparing training runs and querying metadata, with a Python SDK and an Apache 2.0 license.

6.3KUpdated 9 months agoApache-2.0

Docker · Web

An open-source AI music generator that runs locally on macOS, Windows and Linux, with text or audio style prompts and Apache 2.0 code and DiT weights.

2.3KUpdated 10 months agoApache-2.0

macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input

An open-source AI coding assistant for VS Code using Ollama, llama.cpp, LM Studio or hosted APIs, with an MIT-licensed self-hosted team gateway.

3.7KUpdated 1 day agoMIT

Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend

Trilium Notes organizes a personal knowledge base with self-hosted sync and built-in AI chat using local Ollama, LM Studio or cloud providers.

38.1KUpdated 1 day agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Code execution#MCP#Single sign-on

Jina’s embedding models encode multilingual text and media for retrieval, with local weights, noncommercial licenses and commercial deployment options.

jina.aiEmbedding and Reranker Models

Docker#GGUF#LoRA#MLX

Self-hosted Telegram bot for chatting with local LLMs through Ollama. Open source under MIT; archived and no longer maintained.

424Updated 7 months agoMIT

Docker#Multi-user access#Ollama integration

A self-hosted Matrix chatbot using OpenAI APIs, with encrypted-room support and conversation context. Archived and unmaintained.

241Updated 2 years agoAGPL-3.0

Docker#Multi-user access#OpenAI-compatible API

Open-source model optimization library under Apache 2.0. Compress Hugging Face, PyTorch and ONNX models for TensorRT-LLM, vLLM and SGLang.

5.1KUpdated 22 hours agoApache-2.0

Windows · Docker#Agent Skills#Hugging Face integration#ONNX

Open-source AI training framework built on PyTorch. Train on local CPUs or GPUs, fine-tune HuggingFace models, and serve models on your own server.

11.8KUpdated 4 days agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

An open-source Python library for model quantization on your own hardware, with PyTorch, TensorFlow and JAX support under Apache 2.0.

2.7KUpdated 1 week agoApache-2.0

Linux · Docker#Hugging Face integration#Quantization

An open-source LLM fine-tuning tool with a browser interface. Runs on Ubuntu with NVIDIA GPUs or in Docker, under the Apache 2.0 license.

5.2KUpdated 4 days agoApache-2.0

Linux · Docker · Web#Hugging Face integration#LoRA#Quantization

Open-source text-to-speech model with voice cloning. Runs locally on Linux and macOS under Apache 2.0, with a hosted audio playground also available.

7.2KUpdated 2 years agoApache-2.0

macOS · Linux · Docker · Web#Multilingual#Voice cloning

An open source AI character interface you can run locally, with voice chat, VRM avatars, and support for Ollama, llama.cpp, and cloud APIs.

1.6KUpdated 1 year agoMIT

Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input

A self-hosted LLM evaluation tool with a local dashboard, custom checks and root cause analysis. Apache 2.0 licensed; model grading can call cloud APIs.

2.4KUpdated 2 years agoApache-2.0

Docker · Web#Hugging Face integration#Ollama integration

Open-source LLM observability software with self-hosting via Docker, OpenTelemetry traces, and Python and TypeScript SDKs.

1.2KUpdated 10 months agoAGPL-3.0

Docker · Web#LLM tracing#Ollama integration

A self-hosted document AI extension for Paperless-ngx that classifies files and searches your archive with Ollama or cloud APIs. MIT licensed.

6KUpdated 6 months agoMIT

Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API

Favicon of XTTS v2

XTTS v2

1 video
Local text-to-speech model with voice cloning and streaming audio, available through Coqui TTS on Linux, macOS and Windows.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

An open-source AI terminal assistant for Linux, macOS and Windows that uses OpenAI's API or local models through Ollama.

12.3KUpdated 3 months agoMIT

macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval

An open-source AI roleplay chat app you can self-host with Docker, with custom characters, group chats and connections to Kobold, Claude and OpenRouter.

784Updated 4 months agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Multi-user access#Multimodal input#Persistent memory

Open-source AI roleplay client for desktop, Android and web, with Docker hosting and support for OpenAI, Claude, Gemini and Ooba backends.

1.7KUpdated 2 days agoGPL-3.0

macOS · Windows · Linux · Android · Docker · Web#Multilingual#Persistent memory

Self-hosted LLM observability platform for agent tracing, chatbot analytics and prompt management, with Docker, Kubernetes and a cloud offering.

lunary.aiLLM Evaluation and Testing

Docker · Web#LLM tracing#Multi-user access#Prompt versioning

Self-hosted LLM gateway with evaluation, A/B testing, and Ollama support. Open source under Apache 2.0; archived and no longer maintained.

11.7KUpdated 4 months agoApache-2.0

Docker · Web#Batch processing#LLM tracing#Multimodal input

A self-hosted Python library for PostgreSQL RAG and semantic search, with Ollama and cloud embedding providers. Open source, archived and no longer maintained.

5.8KUpdated 4 months agoPostgreSQL

Docker#Batch processing#Ollama integration#RAG

Self-hosted RAG framework for document Q&A, with Docker, Ollama and Infinity support. Apache 2.0 licensed; archived and no longer maintained.

4.4KUpdated 7 months agoApache-2.0

Docker · Web#Batch processing#Multimodal input#Ollama integration

AI agent platform for recurring work, with a hosted service or Docker self-hosting on your own infrastructure and model access.

187.6KUpdated 3 days ago

Docker · Web#Human approval#Multi-agent workflows#Scheduled tasks

An open-source AI agent framework that runs locally in Docker, supports local LLMs with a GPU, and provides a browser interface for managing agents.

17.7KUpdated 2 years agoMIT

Docker · Web#Human approval#Persistent memory#Tool calling

A self-hosted LLM chat interface that runs models through llama.cpp in Docker without remote API keys. Source is licensed under MIT and Apache 2.0.

5.7KUpdated 1 year agoApache-2.0

Windows · Docker · Web#llama.cpp backend

A self-hosted ChatGPT alternative that runs Llama 2 and Code Llama locally, with an MIT license and an OpenAI-compatible API.

10.9KUpdated 3 years agoMIT

macOS · Docker · Web#GGUF#llama.cpp backend#OpenAI-compatible API

Local LLM web interface for Windows, macOS and Linux. Use GGUF models, Ollama or cloud APIs, with local chat storage and an Apache 2.0 license.

4.8KUpdated 3 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

Multimodal AI models you can run locally with Mistral's GPU inference library, which is open source under Apache 2.0 and archived.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#Tool calling