Local & Self-Hosted AI Tools

Browse local and self-hosted AI software, from chat apps and model servers to coding, image, video and voice tools.

800+ tools
An open-source AI app generator you can self-host, with browser previews and cloud inference through Together AI. Licensed under MIT.

7.1KUpdated 2 weeks agoMIT

Web#Code execution

Favicon of Flowise

Flowise

2 videos
Self-hosted AI agent builder with visual workflows, knowledge retrieval and human review. Runs locally or in Docker; the repository is archived.

55.5KUpdated 2 months ago

Docker · Web#Human approval#LLM tracing#Multi-agent workflows

An open-source AI UI builder that runs locally with Docker, previews generated interfaces, and works with Ollama or cloud models.

22.6KUpdated 3 months agoApache-2.0

macOS · Docker · Web#Ollama integration#OpenAI-compatible API

A self-hosted AI coding assistant that builds apps with developer review. Runs locally, uses cloud model APIs, and is no longer maintained.

33.7KUpdated 4 months ago

Windows · Docker#Code execution#Human approval#Multi-agent workflows

A self-hostable model-sharing platform with accounts, uploads and comments. Its public website adds hosted generation that local setup does not include.

7.3KUpdated 17 hours agoApache-2.0

Linux · Docker · Web#Multi-user access

A coding assistant CLI that writes and runs code from plain-language requests. Runs locally or in Docker, with local models or OpenAI and Anthropic APIs.

55.1KUpdated 2 years agoMIT

Windows · Docker#Code execution#Multimodal input

A C++ neural translation library based on Marian NMT, with native and WebAssembly builds for local processing. Open source under MPL 2.0.

550Updated 2 years agoMPL-2.0

Web#Multilingual

An open-source AI email app you can self-host, with Gmail and Outlook connections, a unified inbox, and an MIT license.

10.8KUpdated 1 year agoMIT

Web

AI coding assistant for VS Code, JetBrains and Visual Studio. Connect local models through Ollama or LM Studio, or use cloud models with your own API keys.

codegpt.coCode Completion and IDE Extensions

VS Code · JetBrains#Code execution#Human approval#LM Studio integration

AI content moderation models check text and images in LLM inputs and outputs. Use downloadable models or Meta's hosted Llama API.

4.4KUpdated 20 hours ago

#Guardrails#Multilingual#Multimodal input

An open-source screen history tool that runs offline on Windows, macOS and Linux, with local AI search and screenshots stored on your device.

2.9KUpdated 1 year agoAGPL-3.0

macOS · Windows · Linux · Web#Semantic search#Works offline

An open-source AI code editor built on VS Code that connects to local models or cloud providers directly. Apache 2.0 licensed; archived and unmaintained.

28.8KUpdated 4 months agoApache-2.0

An open source AI code editor built on VS Code, with codebase chat, automated coding and a separate cloud model router. The app repository is MIT licensed.

713Updated 1 year agoMIT

#RAG

An open-source AI coding agent for VS Code and JetBrains, with on-premise deployment and repository-aware assistance. The original project is archived.

3.5KUpdated 4 months agoBSD-3-Clause

VS Code · JetBrains#Code execution#Persistent memory#RAG

AI writing assistant for macOS that works with Apple Foundation Models, Ollama and LM Studio, or cloud providers through your own API keys.

kerlig.comChat With Your Documents

macOS#LM Studio integration#MCP#Multimodal input

A self-hosted AI coding agent that plans tasks, researches the web and writes code, with Ollama support and an MIT license.

19.6KUpdated 1 year agoMIT

macOS · Windows · Linux · Docker · Web#Ollama integration#Web search

A local AI writing app for Windows, Linux and macOS. Keep documents offline, choose your own model, and export in open formats. Open source under AGPL-3.0.

278Updated 1 day agoAGPL-3.0

macOS · Windows · Linux · Android · Web#Works offline

An AI writing workspace for Mac, Windows, iOS and Android. Keep Markdown files locally and use offline AI through Ollama or Apple Intelligence.

hillnote.comAI Notes and Knowledge Bases

macOS · Windows · iOS · Android#MCP#Ollama integration#Works offline

An open-source AI agent platform with local execution, Ollama and LM Studio support, persistent sessions, and a macOS desktop companion. MIT licensed.

35Updated 7 months agoMIT

macOS · Web · VS Code#Code execution#LM Studio integration#MCP

An image-to-3D AI model you can run locally on a GPU or CPU, with GLB export, ComfyUI support and hosted access through Stability AI.

1.8KUpdated 2 years ago

macOS · Windows · Linux#Batch processing#Hugging Face integration

An open-source image-to-3D tool that runs locally with CUDA, exports OBJ meshes, and includes a Gradio interface and Docker support. Apache 2.0 licensed.

4.5KUpdated 2 years agoApache-2.0

Docker · Web#Image-to-image

A local AI audio generator that turns text into sound effects, music and speech. Runs on CPU, NVIDIA CUDA or Apple Silicon with Hugging Face Diffusers support.

2.6KUpdated 2 years ago

macOS · Linux · Web#Batch processing#Hugging Face integration

An open-source image-to-3D model that runs locally with Python and PyTorch. MIT-licensed code and weights, with about 6GB VRAM for default inference.

7KUpdated 4 months agoMIT

#Batch processing

A local text embedding model for semantic search and RAG, with adjustable vector sizes, Apache 2.0 licensing, and support for Sentence Transformers.

1.9KUpdated 11 months ago

Docker#Batch processing#Hugging Face integration#ONNX

An open source LLM programming language that combines Python logic with output constraints and supports local llama.cpp and Transformers models or cloud APIs.

4.2KUpdated 1 year agoApache-2.0

Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration

A paid Windows AI agent app that coordinates custom agents, stores data locally, and supports Ollama alongside ChatGPT and Mistral AI.

nurgo-software.comAI Workflow Automation

Windows#Code execution#Multi-agent workflows#Multimodal input

A local AI coding assistant for VS Code that uses Ollama on your computer or a server you control. Open source under MIT, with no telemetry.

2.1KUpdated 2 years agoMIT

macOS · Windows · VS Code#Multilingual#Ollama integration#Quantization

Open-source data version control for ML projects on macOS, Windows and Linux, with local experiment tracking and a VS Code extension.

15.9KUpdated 2 months agoApache-2.0

macOS · Windows · Linux · VS Code#Git integration

Self-hosted ML experiment tracker for comparing training runs and querying metadata, with a Python SDK and an Apache 2.0 license.

6.3KUpdated 9 months agoApache-2.0

Docker · Web

An open-source Python library for preparing LLM training data locally or on Slurm and Ray clusters, with filtering, deduplication and generation.

3.4KUpdated 1 day agoApache-2.0

#Batch processing#Distributed execution#Multilingual

An open-source AI music generator that runs locally on macOS, Windows and Linux, with text or audio style prompts and Apache 2.0 code and DiT weights.

2.3KUpdated 10 months agoApache-2.0

macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input

An open-source AI coding assistant for VS Code using Ollama, llama.cpp, LM Studio or hosted APIs, with an MIT-licensed self-hosted team gateway.

3.7KUpdated 1 day agoMIT

Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend

An open-source local LLM training toolkit in Python for training GPTs from scratch or fine-tuning GPT-2, with CPU, NVIDIA and Apple Silicon support.

63.5KUpdated 11 months agoMIT

macOS · Windows#Hugging Face integration

An archived, open-source AI coding assistant derived from Cline, with coding, planning and debugging modes. Licensed under Apache 2.0.

24.3KUpdated 5 months agoApache-2.0

VS Code#MCP#Tool calling

A local text-to-audio model for sound effects and music experiments, with CPU or CUDA support and access under the Stability AI Community License.

3.9KUpdated 4 months agoMIT

#Batch processing#Hugging Face integration

A multilingual text embedding model that runs offline on phones, laptops and tablets, with open weights and a quantized memory footprint under 200MB.

5.8KUpdated 11 hours agoApache-2.0

#Hugging Face integration#Multilingual#Quantization

Text embedding models for semantic search, licensed under Apache 2.0, with compact, large and long-context variants for document retrieval.

91Updated 2 years agoApache-2.0

#Hugging Face integration

Alibaba’s GTE models turn text into vectors for retrieval and similarity matching, with downloadable weights and local Python inference.

huggingface.coEmbedding and Reranker Models

#Batch processing#Hugging Face integration#Multilingual

Mixedbread’s mxbai-embed models produce text vectors locally for document retrieval, with Apache-2.0 weights and adjustable embedding dimensions.

mixedbread.comEmbedding and Reranker Models

#Hugging Face integration

Open-source text encoder models under Apache 2.0 for local retrieval and classification, with long context and Hugging Face Transformers support.

1.8KUpdated 7 months agoApache-2.0

#Hugging Face integration

A local document retrieval library that matches text queries to page images without OCR. MIT-licensed Python code supports NVIDIA and Apple Silicon GPUs.

2.8KUpdated 1 month agoMIT

macOS#Batch processing#Hugging Face integration#LoRA

Trilium Notes organizes a personal knowledge base with self-hosted sync and built-in AI chat using local Ollama, LM Studio or cloud providers.

38.1KUpdated 1 day agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Code execution#MCP#Single sign-on

Jina’s embedding models encode multilingual text and media for retrieval, with local weights, noncommercial licenses and commercial deployment options.

jina.aiEmbedding and Reranker Models

Docker#GGUF#LoRA#MLX

Self-hosted Telegram bot for chatting with local LLMs through Ollama. Open source under MIT; archived and no longer maintained.

424Updated 7 months agoMIT

Docker#Multi-user access#Ollama integration

Discourse’s bundled AI plugin adds forum assistants, semantic search, summaries and moderation. Configure a self-hosted model service or cloud provider.

47.9KUpdated 1 day agoGPL-2.0

Linux · Web#LLM tracing#Multi-user access#Structured output

An AI companion that talks, edits files and plays out stories. Run models on your own hardware or use cloud services, with Windows and Linux support.

voxta.aiAI Characters and Roleplay

Windows · Linux · Android · Web#Code execution#MCP#Multimodal input

A self-hosted Matrix chatbot using OpenAI APIs, with encrypted-room support and conversation context. Archived and unmaintained.

241Updated 2 years agoAGPL-3.0

Docker#Multi-user access#OpenAI-compatible API

Open-source model optimization library under Apache 2.0. Compress Hugging Face, PyTorch and ONNX models for TensorRT-LLM, vLLM and SGLang.

5.1KUpdated 22 hours agoApache-2.0

Windows · Docker#Agent Skills#Hugging Face integration#ONNX