Local & Self-Hosted AI Tools

Browse local and self-hosted AI software, from chat apps and model servers to coding, image, video and voice tools.

800+ tools
An open-source visual editor for React websites. Run it locally or self-host it, with AI chat and support for Next.js and Tailwind CSS.

26.8KUpdated 2 months agoApache-2.0

Web#Code execution#Git integration#Multimodal input

Open-source Java library for LLM apps on the JVM, with Ollama support, tool calling, RAG, and integrations for Spring Boot and Quarkus.

13.2KUpdated 1 day agoApache-2.0

#MCP#Ollama integration#RAG

An open-source Ruby AI framework for Ruby and Rails apps, with Ollama, hosted providers and OpenAI-compatible endpoints under an MIT license.

4.4KUpdated 1 day agoMIT

#Human approval#Multi-agent workflows#Multimodal input

An MIT licensed tool that packages local or remote repositories for Claude, ChatGPT, Gemini, and MCP assistants through a command line or website.

28.6KUpdated 2 days agoMIT

Docker · Web · Browser Extension#Agent Skills#Code execution#MCP

A self-hosted AI workspace with parallel model chats and answer merging. Connect Ollama, LM Studio or cloud providers using your own API keys.

7.1KUpdated 1 day agoMIT

Docker · Web#LM Studio integration#Multimodal input#Ollama integration

An open source 3B local LLM that runs on CPU or CUDA GPU and supports direct or reasoning responses, six languages and a 128k context window.

3.9KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Tool calling

An open-source AI data annotation tool for human feedback, fine-tuning and evaluation. Run your own server or deploy on Hugging Face Spaces.

5.1KUpdated 1 year agoApache-2.0

Web#Multi-user access#Semantic search

An open-source OCR toolkit that converts PDFs and images into Markdown using a local GPU or an OpenAI-compatible inference server. Apache 2.0 licensed.

19.7KUpdated 6 months agoApache-2.0

Linux · Docker · Web#Batch processing#Distributed execution#OpenAI-compatible API

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

Open-source meeting bot API for Meet, Teams and Zoom, with live transcripts and AI agent access. Self-host with Docker or use the hosted service.

2.8KUpdated 2 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#Git integration#Human approval

Favicon of PEFT

PEFT

1 video
An open-source Python library for adapting LLMs and diffusion models with less memory and storage, under Apache 2.0.

21.7KUpdated 1 day agoApache-2.0

macOS#Distributed execution#Hugging Face integration#LoRA

.NET library for local LLM apps using Ollama, with streaming chat, embeddings and model management. Open source under MIT; also supports Ollama cloud.

1.4KUpdated 2 months agoMIT

#Multimodal input#Ollama integration#Streaming inference

An open-source segmentation model for images and videos. Run it on your own GPU machine, track objects across frames, and refine masks with prompts.

19.9KUpdated 2 years agoApache-2.0

Web#Hugging Face integration

Open-source AI image and video upscaler with local Python and portable Windows, macOS and Linux versions, licensed under BSD 3-Clause.

36.9KUpdated 2 years agoBSD-3-Clause

macOS · Windows · Linux#Batch processing

An open-source AI frame interpolation tool that processes video locally, builds on RIFE and SAFA, and supports Apple Silicon acceleration under an MIT license.

1KUpdated 1 month agoMIT

macOS

An MIT-licensed LLM data extraction library that validates structured outputs with Pydantic and works with Ollama, llama-cpp-python, vLLM and cloud APIs.

14KUpdated 3 weeks agoMIT

#llama.cpp backend#Ollama integration#Streaming inference

An open-source AI coding assistant for Neovim that connects to Ollama, llama-cpp or cloud providers and applies suggested edits directly to files.

18.2KUpdated 4 days agoApache-2.0

Windows#Agent Client Protocol#llama.cpp backend#Ollama integration

A self-hosted AI agent workspace with Ollama support, shared files, and approval controls. Run it on your servers or use the hosted cloud service.

4.8KUpdated 1 day ago

Docker · Web#Git integration#Human approval#LLM tracing

Self-hosted AI agent platform with document Q&A, workflows and MCP tools. Connect local DeepSeek, Llama or Qwen models, or use cloud providers.

22.9KUpdated 3 days agoGPL-3.0

Docker · Web#MCP#Multimodal input#RAG

Self-hosted speech-to-text API that runs in Docker on CPU or CUDA GPUs, with Whisper, Faster Whisper and WhisperX. Open source under MIT.

3.3KUpdated 2 months agoMIT

Docker · Web#Multilingual#Speaker diarization#Voice activity detection

Favicon of Outlines

Outlines

2 videos
An open-source Python library that constrains LLM outputs to schemas and grammars, with support for local backends including Ollama and llama.cpp.

15.9KUpdated 1 month agoApache-2.0

#llama.cpp backend#MLX#Ollama integration

Local LLM library runs GGUF models through llama.cpp in Node.js, Bun and Electron. MIT licensed, with GPU support and JSON schema enforcement.

2.2KUpdated 3 days agoMIT

macOS · Windows · Linux#Batch processing#GGUF#Guardrails

An open-source dictation app for browser and desktop use, with a Chrome extension and Groq-hosted Whisper transcription.

4.8KUpdated 2 days ago

Web · Browser Extension

Favicon of Scrapling

Scrapling

1 video
Python web scraping framework that runs locally or in Docker, with adaptive parsing, browser automation and an MCP server for AI agents. BSD-3-Clause licensed.

84.5KUpdated 1 day agoBSD-3-Clause

Docker#Guardrails#MCP#Tool calling

An open source AI platform for Kubernetes that supports self-hosted ML workflows, distributed training and model management under Apache 2.0.

15.9KUpdated 1 month agoApache-2.0

Web#MLX#Multi-user access

A local AI image processing editor for Windows, macOS, and Linux. Build visual workflows with community upscaling models. Licensed under GPL-3.0.

6KUpdated 2 months agoGPL-3.0

macOS · Windows · Linux#Batch processing#ONNX#Visual workflows

Self-hosted AI workflow automation with agents, tools, memory and service channels. Source-available under the Fair Core License 1.0.

1.3KUpdated 1 month ago

Web#Human approval#MCP#Multi-user access

An open-source programming language for AI agents that runs on macOS, Linux and Windows, with typed model calls, local tracing and built-in evaluations.

9.4KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · iOS · Android · Web#Batch processing#LLM tracing#Structured output

A self-hosted AI coding assistant for full-stack web apps, with Ollama and LM Studio support, desktop apps, and MIT-licensed source code.

19.9KUpdated 8 months agoMIT

macOS · Windows · Linux · Docker · Web#Code execution#Git integration#LM Studio integration

An MIT-licensed translation app for Windows, macOS and Linux that runs Marian and Bergamot models locally and connects to Chrome and Firefox.

631Updated 2 years agoMIT

macOS · Windows · Linux · Web · Browser Extension#Batch processing#Multilingual#Quantization

Favicon of Pinokio

Pinokio

5 videos
A local AI app launcher for Windows, macOS, and Linux that helps you find and run community projects. It's open source under the MIT license.

8.2KUpdated 4 weeks agoMIT

macOS · Windows · Linux#Code execution

An open source Neovim AI coding assistant with Ollama support, cloud model connections, inline edits, and agents such as Claude Code and Codex.

6.9KUpdated 22 hours agoApache-2.0

#Agent Client Protocol#Human approval#MCP

Open-source Python AI agent framework under MIT. Coordinate agents and tools, use local or hosted sandboxes, and connect to OpenAI or other model providers.

29.8KUpdated 1 day agoMIT

macOS · Windows · Linux#Code execution#Guardrails#Human approval

Favicon of Axolotl

Axolotl

1 video
Apache 2.0 LLM fine-tuning framework for local or cloud GPUs, supporting NVIDIA, AMD, LoRA, QLoRA and multimodal training.

12.5KUpdated 2 days agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

An offline speech-to-text and text-to-speech app for Linux and Sailfish OS, with local translation, voice typing and MPL-2.0 open-source licensing.

1.7KUpdated 1 week agoMPL-2.0

Linux#Multilingual#Works offline

An open-source NVIDIA GPU process monitor for Linux and Windows, with interactive terminal views, Python APIs and Grafana dashboard support.

7.2KUpdated 2 days agoApache-2.0

Windows · Linux · Docker · Web

Open-source speaker diarization toolkit with local PyTorch models, CUDA GPU support, and an optional hosted service that processes audio on pyannoteAI servers.

10.6KUpdated 3 months agoMIT

#Hugging Face integration#Speaker diarization#Voice activity detection

An open-source AI chat app for macOS, iOS and visionOS that connects to your own Ollama server and stores conversation history on your device.

6KUpdated 3 months agoApache-2.0

macOS · iOS#Multimodal input#Ollama integration#Works offline

Open-source text-to-speech toolkit for local speech generation, voice cloning and model training on Linux, macOS and Windows, licensed under MPL-2.0.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

An open-source LLM client that brings local models through Ollama and llama.cpp, plus cloud services such as Claude and Gemini, into Emacs.

3.5KUpdated 6 days agoGPL-3.0

#Git integration#Human approval#llama.cpp backend

Self-hosted AI observability platform for tracing and evaluating LLM apps. Run it locally or in Docker, with Ollama integration and OpenTelemetry support.

11.7KUpdated 1 day ago

Docker · Web#LLM tracing#MCP#Ollama integration

A self-hosted inference framework that coordinates NVIDIA GPU clusters with vLLM, SGLang or TensorRT-LLM and exposes an OpenAI-compatible API.

8.2KUpdated 21 hours ago

#Distributed execution#Multimodal input#OpenAI-compatible API

A self-hosted OCR model that extracts text and page structure from PDFs and images, with vLLM, Hugging Face Transformers and CPU inference support.

9.2KUpdated 6 months agoMIT

Docker#Hugging Face integration#Multilingual#Multimodal input

Open-source Python toolkit for semantic search and RAG, with BGE embedding models, multilingual rerankers, evaluation and fine-tuning under MIT.

12.2KUpdated 1 month agoMIT

#Multilingual#Semantic search

A Python research assistant that answers questions with citations from local documents. Supports self-hosted models through LiteLLM; licensed under Apache 2.0.

9.3KUpdated 2 months agoApache-2.0

#Multilingual#Multimodal input#RAG

Favicon of OpenVINO

OpenVINO

2 videos
Open-source AI inference toolkit for local or self-hosted deployment on Linux, Windows and macOS, with CPU, Intel GPU and NPU support.

10.9KUpdated 1 day agoApache-2.0

macOS · Windows · Linux#Hugging Face integration#Multimodal input#ONNX

Favicon of Milvus

Milvus

2 videos
An open-source vector database for AI retrieval. Run it locally or on your servers, with hybrid search, metadata filtering and CPU or GPU acceleration.

46.3KUpdated 1 day agoApache-2.0

macOS · Linux#Hybrid search#Semantic search

An open-source Postgres search extension for full-text and vector retrieval, filters and aggregations. Run it locally or self-host it under AGPL-3.0.

9.3KUpdated 21 hours agoAGPL-3.0

Docker#MCP#Semantic search