Local & Self-Hosted AI Tools

Browse local and self-hosted AI software, from chat apps and model servers to coding, image, video and voice tools.

800+ tools
AI agent platform for recurring work, with a hosted service or Docker self-hosting on your own infrastructure and model access.

187.6KUpdated 3 days ago

Docker · Web#Human approval#Multi-agent workflows#Scheduled tasks

Favicon of AutoGen

AutoGen

2 videos
AI agent framework for Python with local and distributed runtimes, a browser-based prototyping UI, and OpenAI and Azure OpenAI integrations.

61.2KUpdated 6 months agoCC-BY-4.0

Web#Code execution#MCP#Multi-agent workflows

An open-source AI agent framework that runs locally in Docker, supports local LLMs with a GPU, and provides a browser interface for managing agents.

17.7KUpdated 2 years agoMIT

Docker · Web#Human approval#Persistent memory#Tool calling

Self-hosted AI application server with an OpenAI-compatible API, local Ollama and vLLM backends, document search and agent tool calling. MIT licensed.

8.4KUpdated 22 hours agoMIT

#Agent Skills#Batch processing#Guardrails

A self-hosted LLM chat interface that runs models through llama.cpp in Docker without remote API keys. Source is licensed under MIT and Apache 2.0.

5.7KUpdated 1 year agoApache-2.0

Windows · Docker · Web#llama.cpp backend

A self-hosted ChatGPT alternative that runs Llama 2 and Code Llama locally, with an MIT license and an OpenAI-compatible API.

10.9KUpdated 3 years agoMIT

macOS · Docker · Web#GGUF#llama.cpp backend#OpenAI-compatible API

An offline AI assistant for Android and iOS that runs models on your phone, keeps chats encrypted on-device, and offers optional hosted models.

layla-network.aiAI Characters and Roleplay

iOS · Android#Code execution#GGUF#llama.cpp backend

Local LLM web interface for Windows, macOS and Linux. Use GGUF models, Ollama or cloud APIs, with local chat storage and an Apache 2.0 license.

4.8KUpdated 3 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend

A self-hosted AI chat app that connects to Ollama and cloud model APIs, with an MIT license and a Supabase backend you can run locally.

33.3KUpdated 2 years agoMIT

macOS · Windows · Linux · Web#Ollama integration

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An image generation model you can run locally, with text-to-image and image-to-image support, MIT-licensed Python code and separately licensed weights.

27.3KUpdated 9 months agoMIT

Web#Hugging Face integration#Image-to-image

Favicon of Qwen-Image

Qwen-Image

1 video
An open-source image generation and editing model for local deployment, with Chinese text rendering, multi-image edits and an Apache 2.0 license.

8.4KUpdated 8 months agoApache-2.0

Web#Image-to-image#LoRA#Multimodal input

A local AI image generation model with text-guided image editing, Diffusers support, and a reference implementation requiring at least 10GB of GPU VRAM.

73.5KUpdated 4 years ago

#Guardrails#Hugging Face integration#Image-to-image

An on-device AI SDK that runs text, image and audio models on macOS, Windows and Linux, with GGUF, MLX and an OpenAI-compatible API.

qualcomm/GenieXInference Libraries and Bindings

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

A local vision-language model for image analysis, document extraction and video understanding, with Apache 2.0 weights and Hugging Face Transformers support.

huggingface.coOCR and Document Scanning

Linux#Batch processing#Hugging Face integration#Multimodal input

Multimodal AI models you can run locally with Mistral's GPU inference library, which is open source under Apache 2.0 and archived.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#Tool calling

Open-source vision language model for local image and text tasks, with Apache 2.0 licensing, Transformers support, and a small GPU memory footprint.

3.9KUpdated 1 week agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

An open source on-device AI engine for local LLMs and image models, with iOS, Android, CPU and GPU support under Apache 2.0.

16.2KUpdated 1 day agoApache-2.0

Windows · iOS · Android#Image-to-image#Multimodal input#ONNX

Self-hosted AI data assistant that queries databases, analyzes files, and generates reports with local models or cloud APIs. Open source under MIT.

20.1KUpdated 2 days agoMIT

macOS · Linux · Docker · Web#Code execution#llama.cpp backend#OpenAI-compatible API

Self-hosted AI agent framework in C# for Windows, Linux and macOS, with LLamaSharp and cloud provider plugins. Apache 2.0 licensed.

3.1KUpdated 2 days agoApache-2.0

macOS · Windows · Linux · Web#Code execution#MCP#Multi-agent workflows

Self-hosted photo management software with face recognition and semantic search. Runs through Docker, supports RAW photos and videos, and uses the MIT license.

8.1KUpdated 1 day agoMIT

Windows · Linux · iOS · Android · Docker · Web#Multi-user access#ONNX#Semantic search

Open-source coding LLMs for code generation, reasoning and fixes. Run the 32B instruction model locally with Transformers or self-host it with vLLM.

huggingface.coCoding Models

#Hugging Face integration

Free, open source CCTV software for Linux that captures, analyzes and records video from IP, USB and analog cameras under the GPL-2.0 license.

5.9KUpdated 22 hours agoGPL-2.0

Linux · Docker

A self-hosted LLM with a 256K context window, vLLM and Transformers support, and research and commercial use under the Jamba Open Model License.

huggingface.coOpen-Weight LLMs

#Hugging Face integration#LoRA#Multilingual

A self-hosted Kubernetes operator that deploys Hugging Face models with vLLM, manages GPU capacity, and runs fine-tuning and document retrieval services.

1KUpdated 1 week ago

#Distributed execution#Hugging Face integration#LoRA

Downloadable coding models for local completion and tool-using agents. Codestral and Devstral releases have different model licenses.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Tool calling

An open weights LLM you can self-host with Transformers, with document citations, tool calling and a CC-BY-NC-4.0 noncommercial license.

huggingface.coOpen-Weight LLMs

#Guardrails#Hugging Face integration#Multilingual

A self-hosted AI serving framework that deploys models and pipelines on Kubernetes, on-premises or in the cloud, under the Business Source License.

4.8KUpdated 8 months ago

A family of AI models you can run offline with Ollama, llama.cpp or LM Studio, with open weights and training data for building specialized agents.

2.1KUpdated 3 weeks agoApache-2.0

Linux#GGUF#Guardrails#Hugging Face integration

LLM models for local or cloud inference, with downloadable weights and function calling. The Apache 2.0 inference library is archived.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#Tool calling

Open-source LLM for self-hosted AI agents, with thinking and direct-response modes, MIT licensing, and support for vLLM and SGLang.

huggingface.coCoding Models

#Hugging Face integration#LoRA#Multilingual

An open-source Go library for building LLM applications with Ollama, OpenAI and Gemini, plus components for agents and document retrieval. MIT licensed.

9.7KUpdated 9 months agoMIT

#Ollama integration#RAG#Semantic search

An open source LLM evaluation framework for local models and hosted APIs, with GGUF, Hugging Face transformers and llama.cpp support. MIT licensed.

14.1KUpdated 2 weeks agoMIT

macOS#Batch processing#GGUF#Hugging Face integration

An open-source AI gateway built on Envoy. Self-host agent orchestration, route LLM requests, and capture OpenTelemetry traces under Apache 2.0.

7.1KUpdated 2 days agoApache-2.0

Docker#Guardrails#LLM tracing#Multi-agent workflows

An open-source portrait animation tool that runs on Ubuntu with an NVIDIA GPU, turning a still image and English speech into a talking video.

8.7KUpdated 2 years agoMIT

Linux#Hugging Face integration#Multimodal input#ONNX

Self-hosted AI image generation UI for Windows, Linux and Apple Silicon Macs. Use Stable Diffusion and Flux with ComfyUI workflows under an MIT license.

4.6KUpdated 22 hours agoMIT

macOS · Windows · Linux · Docker · Web#Visual workflows

An open-source Python document parser for LLM applications. Run it locally or in Docker to process PDFs, Word files and images under Apache 2.0.

15.5KUpdated 3 days agoApache-2.0

macOS · Windows · Linux · Docker#Multilingual

An open-source AI gateway under the MIT license, with an OpenAI-compatible API for Ollama and cloud providers. Runs locally with Node.js or Docker.

13.1KUpdated 4 months agoMIT

Docker · Web#Guardrails#Ollama integration#OpenAI-compatible API

A self-hosted multimodal retrieval engine for AI apps, with ColPali visual search, Docker deployment, and MCP access through Claude or Open Web UI.

3.7KUpdated 5 days ago

Docker · Web#MCP#Multi-user access#Multimodal input

A local AI video interpolation app for Windows that supports RIFE, DAIN and FLAVR, with GPU processing on NVIDIA and AMD hardware. Open source under GPL-3.0.

2.1KUpdated 4 months agoGPL-3.0

Windows

A self-hosted vector search engine for text and images, with built-in embedding generation, Docker deployment, and an Apache 2.0 license.

5KUpdated 6 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#RAG

An open-source vector search library for local apps and servers, with custom metrics, disk-backed indexes, and Apache 2.0 licensing.

4.3KUpdated 4 weeks agoApache-2.0

macOS · Windows · Linux · iOS · Android#Batch processing#Semantic search

Favicon of Amuse

Amuse

1 video
Local AI creation app for Windows with image, video, audio and text models, GGUF and ONNX support, and acceleration for NVIDIA, AMD and Intel GPUs.

606Updated 4 days agoApache-2.0

Windows#ControlNet#GGUF#Inpainting

A free Stable Diffusion animation tool with local Python and Jupyter runtimes, plus cloud options through Colab and Replicate. No longer maintained.

2.3KUpdated 2 years ago

Windows · Web#Hugging Face integration

AI coding agent for multi-file tasks, with a Docker self-hosted server, MIT license, and support for Anthropic, OpenAI and Google models.

15.7KUpdated 12 months agoMIT

Windows · Docker#Code execution#Git integration#Human approval

A self-hostable AI chat interface with Docker deployment, local model connections through LocalAI and RWKV-Runner, and desktop clients.

88.8KUpdated 2 months agoMIT

macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API

A local LLM app that runs models offline on iOS and macOS, supports text and vision models, and uses ggml and llama.cpp under the MIT license.

2.1KUpdated 8 months agoMIT

macOS · iOS#llama.cpp backend#Multimodal input#RAG

A self-hosted AI retrieval system with document search, cited answers and research agents. Runs in Python or Docker and uses the MIT license.

8KUpdated 11 months agoMIT

Docker#Hybrid search#Knowledge graphs#Multi-user access