Local & Self-Hosted AI Tools

Browse local and self-hosted AI software, from chat apps and model servers to coding, image, video and voice tools.

800+ tools
A self-hosted RSS reader that ranks news using AI tags and your scoring rules. Runs with Docker and supports OpenAI, Gemini and compatible model APIs.

395Updated 4 weeks ago

Docker · Web#Multi-user access#OpenAI-compatible API

Favicon of Draw Things

Draw Things

4 videos
An AI image generator that runs offline on iPhone, iPad and Mac, with on-device LoRA training and optional self-hosted or managed cloud compute.

575Updated 1 day agoGPL-3.0

macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA

An open-source Python framework for AI model serving. Build inference APIs and multi-model pipelines locally or deploy with Docker under Apache 2.0.

8.9KUpdated 3 weeks agoApache-2.0

Docker#Batch processing#ControlNet#Distributed execution

Open-source Python toolkit for training and serving LLMs on your own hardware, with Apache 2.0 licensing, quantization and multi-GPU support.

13.7KUpdated 3 weeks agoApache-2.0

#Hugging Face integration#LoRA#Quantization

An open source RLHF framework for training models on your own NVIDIA GPUs, with HuggingFace model support and Ray, vLLM and DeepSpeed backends.

10.1KUpdated 2 weeks agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

Favicon of Dyad

Dyad

3 videos
A desktop AI app builder for macOS and Windows that turns prompts into web apps, works with Ollama and cloud models, and keeps project files on your machine.

21.6KUpdated 3 hours ago

macOS · Windows#LM Studio integration#MCP#Ollama integration

A local LLM frontend that connects to local backends and cloud APIs, with lorebooks, image generation and voice. Open source under AGPL-3.0.

33.9KUpdated 2 weeks agoAGPL-3.0

macOS · Windows · Linux · Android · Docker · Web#Multilingual

Favicon of LanceDB

LanceDB

1 video
An open source vector database that runs locally or in your cloud, with multimodal storage, hybrid search and Python, TypeScript and Rust SDKs.

11.6KUpdated 23 hours agoApache-2.0

#Hybrid search#Semantic search

Favicon of Inbox Zero

Inbox Zero

1 video
AI email assistant with hosted and self-hosted options. Organize Gmail and Outlook, draft replies in your voice, and set email rules in plain English.

12.4KUpdated 23 hours ago

iOS · Android · Docker · Web#Batch processing

AI gateway and LLM observability platform with Apache 2.0 self-hosting, Ollama integration, request tracing, and cost tracking.

6.2KUpdated 2 weeks agoApache-2.0

Web#LLM tracing#Ollama integration#Prompt versioning

An open-source LLM evaluation framework in Python, licensed under MIT, with local inference through Hugging Face, vLLM and SGLang.

2.9KUpdated 24 hours agoMIT

Web · VS Code#Code execution#Hugging Face integration#MCP

Command-line client for the Hugging Face Hub. Download models to a local cache, manage repositories, and access hosted inference. Open source under Apache 2.0.

3.9KUpdated 1 day agoApache-2.0

#Hugging Face integration

Favicon of LatentSync

LatentSync

1 video
Open-source local AI lip-sync tool with a Gradio interface and Apache 2.0 code license. GPU inference requires 8 GB or 18 GB VRAM, depending on the model.

6.1KUpdated 1 year agoApache-2.0

Web#Batch processing#Hugging Face integration#Multimodal input

A local AI portrait animation tool that transfers facial motion from video to images, with eye and lip controls. Runs on Linux, Windows and Apple Silicon.

19.1KUpdated 4 months ago

macOS · Windows · Linux · Web#Hugging Face integration

An open-source local LLM compiler and deployment engine with GPU support across desktop, browser and mobile platforms, plus an OpenAI-compatible API.

23.2KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · iOS · Android · Web#OpenAI-compatible API

Open-source browser agent SDK for local Chrome or Browserbase cloud browsers, with natural-language actions and structured extraction. MIT licensed.

25.5KUpdated 2 days agoMIT

#LLM tracing#Structured output

A Python library that constrains LLM output with grammars and JSON schemas, supporting local Transformers and llama.cpp backends plus OpenAI.

21.8KUpdated 4 months agoMIT

#llama.cpp backend#Structured output#Tool calling

Favicon of Chatbox

Chatbox

1 video
AI chat client that connects to local models through Ollama and cloud models with your own API keys. Runs on desktop, mobile and the web.

41.9KUpdated 6 days agoGPL-3.0

macOS · Windows · Linux · iOS · Android · Web#MCP#Multimodal input#Ollama integration

A local LLM inference engine with OpenAI and Anthropic-compatible APIs. Runs on macOS, Linux and Windows with CPU, CUDA or Apple Silicon support.

7.7KUpdated 5 days agoMIT

macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Hugging Face integration

Self-hosted AI document search and agents with source citations. Run local models through Ollama, vLLM or llama.cpp, including fully offline deployments.

18.3KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend

An open-source PDF translator that preserves formulas and layouts, runs locally or in Docker, and supports Ollama, Google, DeepL and OpenAI.

37.3KUpdated 1 day agoAGPL-3.0

macOS · Windows · Docker · Web#Batch processing#Hugging Face integration#MCP

An open-source LLM vulnerability scanner that tests local Hugging Face and GGUF models or cloud APIs for security failures. Apache 2.0 licensed.

9.4KUpdated 2 weeks agoApache-2.0

#AI red teaming#GGUF#Hugging Face integration

A self-hosted AI assistant for Nextcloud Hub. Use local models to keep data on your server, or connect to cloud providers. Open source under AGPL-3.0.

91Updated 1 day agoAGPL-3.0

Web#Multi-user access#Multilingual#Multimodal input

Open-source PostgreSQL extension for vector search with pgvector, DiskANN indexing and compression. Run it self-hosted or in Timescale Cloud.

3.1KUpdated 3 weeks agoPostgreSQL

macOS · Linux · Docker#Semantic search

Local AI 3D generator that creates meshes, radiance fields and 3D Gaussians from text or images. Runs on Linux with an NVIDIA GPU and 16GB VRAM.

13.7KUpdated 11 months agoMIT

Linux · Web#Hugging Face integration#Multimodal input

An open-source Obsidian AI plugin for searching and writing notes with Ollama, LM Studio, Claude Code or Codex. Local indexes stay on your device.

7.8KUpdated 2 days agoAGPL-3.0

#LM Studio integration#Multi-agent workflows#Ollama integration

An open-source Stable Diffusion app for macOS that generates and edits images offline, with support for SDXL, ControlNet and local model training.

13.6KUpdated 3 years agoAGPL-3.0

macOS#ControlNet#Image-to-image#Inpainting

An open-source LLM fine-tuning library for single-GPU and distributed training, with LoRA and QLoRA support. Development is no longer active.

5.8KUpdated 5 months agoBSD-3-Clause

#Hugging Face integration#LoRA#Quantization

Favicon of Elia

Elia

1 video
A terminal chat client for local LLMs through Ollama or LocalAI and cloud models from OpenAI and Anthropic, with locally stored conversations.

2.5KUpdated 2 years agoApache-2.0

#Ollama integration#OpenAI-compatible API

A self-hosted AI agent development platform with visual workflows, knowledge bases and plugins, licensed under Apache 2.0, with OpenAI integration.

21.7KUpdated 2 months agoApache-2.0

Web#Persistent memory#RAG#Tool calling

A self-hosted text embedding server with a REST API, CPU and GPU support, and offline operation with downloaded model weights. Apache 2.0 licensed.

5.1KUpdated 1 week agoApache-2.0

macOS · Linux · Docker#Batch processing#Hugging Face integration#LLM tracing

An open-source local AI API under Apache 2.0 that connects to Ollama, llama.cpp and other OpenAI-compatible servers for document retrieval and agent workflows.

57.6KUpdated 1 week agoApache-2.0

Docker · Web#Code execution#llama.cpp backend#MCP

An open-weight vision model for image questions, captions and object detection. Run it locally or use hosted inference and fine-tuning.

10.1KUpdated 5 months agoApache-2.0

macOS · Windows · Linux#Hugging Face integration#Multimodal input#Works offline

Self-hosted computer vision server for images and video, with Docker support, NVIDIA GPU acceleration, and optional Roboflow hosted compute.

2.5KUpdated 1 day ago

macOS · Windows · Linux · Docker#Batch processing#Code execution#Multimodal input

An Obsidian plugin that finds related notes using local embeddings. Works offline by default, keeps notes on your device, and needs no API key.

5.5KUpdated 6 days ago

#Semantic search#Works offline

Local audiobook converter turns ebooks into narrated audio with chapters and voice cloning. Runs on Windows, macOS and Linux under Apache 2.0.

20.3KUpdated 4 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning

Favicon of Lemonade

Lemonade

1 video
An open source local AI server for chat, image generation, and speech on Windows, macOS, and Linux, with APIs for apps and agents.

5.8KUpdated 3 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Docker#GGUF#Hugging Face integration#llama.cpp backend

A self-hosted subtitle translator that runs in Docker and automates media library translation using local Ollama models or cloud services such as DeepL.

881Updated 3 days ago

Docker#Multilingual#Ollama integration

Self-hosted workflow platform turns Python, TypeScript and other scripts into APIs and internal apps. Runs on Docker or Kubernetes; cloud hosting is available.

18.1KUpdated 24 hours ago

Linux · Docker · Web#Code execution#Git integration#Human approval

Favicon of Letta

Letta

1 video
An open-source AI agent platform with persistent memory. Run agents locally or on a self-hosted server, with desktop apps for macOS, Windows, and Linux.

25KUpdated 3 weeks agoApache-2.0

macOS · Windows · Linux#Persistent memory

Open-source Python toolkit for local AI audio generation, fine-tuning and training, with a Gradio interface and support for Stable Audio Open.

3.9KUpdated 4 months agoMIT

Web#Hugging Face integration

Favicon of Presidio

Presidio

1 video
A self-hosted Python framework that detects and anonymizes sensitive data in text, images and structured records. Open source under the MIT license.

11.1KUpdated 2 days agoMIT

Docker#Human approval#Multilingual

Self-hosted AI research assistant for web, papers and private documents. Runs on Windows, macOS and Linux with Ollama or cloud models. MIT licensed.

9.1KUpdated 23 hours agoMIT

macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access

Open-source AI framework for semantic search, RAG and agents. Runs locally or in Docker, with Hugging Face, llama.cpp and cloud models via LiteLLM.

13KUpdated 23 hours agoApache-2.0

Docker#Agent Skills#Hugging Face integration#Knowledge graphs

Favicon of YuE

YuE

1 video
An open-source AI music generator that turns lyrics into songs. Run it locally on Linux with an NVIDIA GPU, or use the hosted demo.

10.6KUpdated 2 days agoApache-2.0

Linux · Web#Hugging Face integration#Multilingual#Multimodal input

Favicon of DSPy

DSPy

2 videos
Open-source Python framework for building modular LLM systems, with typed outputs, tool-using agents, and automatic prompt optimization. MIT licensed.

38.4KUpdated 4 days agoMIT

#Code execution#MCP#Multimodal input

Favicon of Chatwoot

Chatwoot

1 video
Self-hosted customer support platform with an AI agent, shared inbox and help center. Run it on your own server or use the hosted cloud service.

37.3KUpdated 21 hours ago

Docker · Web#Multi-user access#Multilingual#RAG

A local LLM family for developers and researchers, with downloadable weights, text and vision models, and custom licensing for research and commercial use.

7.7KUpdated 12 months ago

#Hugging Face integration#Multimodal input#Quantization