4.4KUpdated 7 months agoMIT
Docker#MCP#Tool calling
mcpo makes MCP tools accessible to AI agents and apps that accept OpenAPI servers, including Open WebUI. It's a self-hosted proxy for developers who want to use existing MCP servers with those apps without writing a separate adapter for each tool. You can run it locally or on your own server with Python or Docker. It's open source under the MIT license.
4.7KUpdated 2 days agoMIT
Docker · Web#Git integration#LLM tracing#MCP
Latitude is a self-hosted AI agent monitoring platform for teams that need to find production failures and check that fixes work. It connects recurring problems to the sessions that caused them, then can dispatch Claude Code or Cursor with enough context to make a fix and open a pull request.
49.1KUpdated 2 days agoAGPL-3.0
Linux · Docker · Web#Multi-user access#Multimodal input#OpenAI-compatible API
New API is a self-hosted AI gateway for developers and teams that want several model providers behind one service. It builds on One API and converts between OpenAI Chat Completions, Responses, Anthropic Messages and Gemini formats, so apps and agents can switch providers without changing each client's connection settings.
6.7KUpdated 2 months agoApache-2.0
Windows · Docker · Web#Batch processing#Hugging Face integration#Multilingual
MonkeyOCR is a local AI document parser for developers and researchers working with English and Chinese PDFs or images. It extracts text, formulas and tables while identifying page structure and relationships between blocks. That makes it useful for documents where plain text extraction loses reading order or separates content from its layout.
38.7KUpdated 11 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Langchain-Chatchat is a self-hosted application for asking questions about your own documents and using AI agents. It focuses on Chinese-language use and open models, with a fully offline setup that can keep documents and model processing on your hardware. Its code is open source under Apache 2.0.
41.2KUpdated 1 day agoAGPL-3.0
macOS · Docker · Web#Code execution#Hybrid search#LM Studio integration
AstrBot brings AI assistants into messaging apps such as Telegram, Discord, Slack, QQ and WeCom. It's an open source platform under AGPL-3.0 for people building personal companions, customer support bots or team automation. You can run it on your own computer or server, including through Docker, or use its desktop app for browser-style chat.
1.4KUpdated 2 days agoAGPL-3.0
Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration
TabbyAPI is a self-hosted LLM API server built around ExLlamaV3, for people who want local model inference behind an OpenAI-compatible API. It's the official server for that backend. The project targets personal use and small groups, and its maintainers explicitly advise against using it for production workloads.
7.1KUpdated 2 days agoApache-2.0
Docker#Batch processing#Distributed execution#Multimodal input
Data-Juicer is a Python framework for preparing AI datasets on your own machine or a distributed Ray cluster. It's for researchers and teams curating model training data, agent interaction records or documents for retrieval. The project is open source under Apache 2.0.
11.1KUpdated 20 hours ago
Docker · Web#Multimodal input#Voice activity detection
TEN Framework is a self-hosted framework for developers building voice AI agents and multimodal conversational apps. It focuses on low-latency, real-time conversations and supports both RTC and WebSocket connections. You can run its agent examples locally with Docker or deploy them on your own server.
2.9KUpdated 9 months agoApache-2.0
Windows · Docker · Web#Hugging Face integration#Multilingual#Speaker diarization
Whisper WebUI turns audio into transcripts and subtitles through a browser interface that runs on your own machine or a self-hosted server. It's for people captioning videos, transcribing recordings or translating spoken content who want local speech processing. The project is open source under Apache 2.0 and supports Docker and Pinokio.
22.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
localGPT is a self-hosted AI document chat app for people who want to question and summarise files on their own hardware. Its local Ollama setup keeps documents and conversations on your machine. Answers include source passages, so you can check what the model used.
16.8KUpdated 1 day agoMIT
Docker · Web#Hugging Face integration#Multi-user access#ONNX
CVAT is a browser-based data annotation platform for teams building computer vision datasets. Its open-source Community edition runs on your own infrastructure with Docker and uses the MIT license. CVAT Online is hosted by CVAT, while the Enterprise offering runs in an organization's own cloud or internal environment.
47.7KUpdated 1 month agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#GGUF#llama.cpp backend#LoRA
text-generation-webui, also called TextGen, runs language models on your own hardware through a desktop app or a self-hosted browser interface. It's for people who want private chat and writing tools, and developers who need a local model API. It works offline without telemetry; web search and page fetching use the internet.
5.5KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#Home Assistant integration#Multilingual#OpenAI-compatible API
Kokoro-FastAPI runs the Kokoro-82M speech model on your own machine or server and exposes an OpenAI-compatible speech API. It's for developers adding local text-to-speech to assistants, reading apps or audiobook workflows. Speech generation runs locally, and the API doesn't require an OpenAI account.
15.7KUpdated 1 year agoMIT
Docker · Web
Gitingest turns a Git repository or local directory into a text digest that developers can give to an LLM as code context. It combines the directory tree and file contents in one extract, so you don't have to assemble context file by file. It's open source under the MIT license.
11KUpdated 1 week agoBSD-3-Clause
Windows · Linux · Docker#Batch processing#ONNX
Triton Inference Server, offered by NVIDIA as Dynamo-Triton, is a self-hosted AI inference server for teams deploying models in applications. It serves models from different frameworks through one server, with support for on-premises hardware, cloud infrastructure and edge devices. It's open source under the BSD-3-Clause license.
40.3KUpdated 22 hours ago
macOS · Windows · Linux · Docker · Web
PhotoPrism is a self-hosted photo and video library for people who want to organize personal media on their own hardware or server. Its AI recognizes faces and labels pictures by content and location, so finding a photo doesn't depend entirely on folders or tags you've added yourself. You browse and share the library through a web app.
84.5KUpdated 6 days agoApache-2.0
Docker#Structured output
Crawl4AI is a self-hosted web crawler and scraper for developers building AI agents, retrieval-augmented generation (RAG) systems and data pipelines. It turns web pages into Markdown or structured JSON and runs as a Python library or a Docker server on your own hardware. The open-source code uses the Apache 2.0 license.
19.2KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual
pyVideoTrans translates spoken audio into another language and produces a video with translated subtitles and AI dubbing. It's for people adapting videos for audiences in other languages who want control over which parts run locally. It recognizes speech directly, so the original video doesn't need subtitles.
12.6KUpdated 3 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#ControlNet#Hugging Face integration#LoRA
Kohya's GUI lets you train and fine-tune image generation models on your own GPU-equipped computer through a browser interface. It's for artists and model makers who want to teach a model a particular style or subject while controlling the training settings. The interface builds on Kohya's Stable Diffusion training scripts, with a command-line interface available too.
13.2KUpdated 1 day agoMIT
Linux · Docker#Git integration#Ollama integration
PR-Agent is an open source AI code review agent for development teams that want to choose where their reviewer runs and which model it uses. It reviews pull requests, writes descriptions, suggests code improvements, and answers questions about proposed changes. You can run it locally through a CLI, on a self-hosted server, in Docker, or through GitHub Actions.
10.8KUpdated 8 months agoMIT
Windows · Docker · Web#Multi-user access#Multilingual
Doccano is a self-hosted text annotation tool for machine learning practitioners who need labeled training or evaluation data. It runs on your own machine or server, with a browser interface and Docker support. The software is open source under the MIT license.
21.7KUpdated 23 hours agoApache-2.0
Docker · Web#Human approval#MCP#Multi-agent workflows
Google ADK is an open-source AI agent framework for developers building applications that carry out multi-step tasks. You can run agents locally in Docker or on your own infrastructure, and connect them to locally running models through adapters. The framework is optimized for Gemini but supports other models and providers. Its Python repository uses the Apache 2.0 license.
25.8KUpdated 4 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access
kotaemon is a self-hosted document chat app for people who want to ask questions across their files and check where the answers came from. It runs in a browser on Windows, macOS or Linux, with Docker also supported. The project uses the Apache 2.0 license.
3.1KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker#GGUF#Hugging Face integration#llama.cpp backend
RamaLama runs and serves AI models on your own hardware using OCI containers. It's aimed at developers who want local chat or a self-hosted inference API with a container workflow they can also use in production. The project uses the MIT license.
11.9KUpdated 4 days agoAGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
KoboldCpp pairs local model inference with a browser interface built for chat, creative writing and roleplay. A fork of llama.cpp, it bundles KoboldAI Lite with tools for keeping character details and story context alongside your conversations. It's open source under AGPL-3.0.
390Updated 1 day agoMIT
Docker#Home Assistant integration#Hugging Face integration#Multilingual
Wyoming Faster Whisper is a local speech-to-text server for Home Assistant and other clients that use the Wyoming protocol. It turns spoken audio into text on your own hardware, with support for names specific to your home. It's open source under the MIT license and runs as a Home Assistant add-on, a Docker container, or a local Python service.
26.6KUpdated 6 days agoGPL-3.0
macOS · Linux · Docker#Hybrid search#Multimodal input#RAG
Typesense combines typo-tolerant site search with vector and semantic search in a self-hosted engine. It's for developers building searchable apps, product catalogs or AI search over their own data. The C++ engine uses an in-memory architecture for low-latency results as users type.
1.4KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker#Batch processing#ONNX
python-audio-separator separates recordings into vocals, instrumentals and individual instruments on your own hardware. It's an open-source Python package under the MIT license, aimed at karaoke creators and developers who want audio separation in scripts or their own applications.
395Updated 4 weeks ago
Docker · Web#Multi-user access#OpenAI-compatible API
Feeds Fun is a news reader for people whose RSS subscriptions produce more articles than they want to read. It assigns tags automatically, then uses rules you define to score articles by topic. You can self-host it with Docker or use the hosted service at feeds.fun.
575Updated 1 day agoGPL-3.0
macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA
Draw Things is an AI image generation app for iPhone, iPad and Mac that keeps generation on your device and works offline. It's for people who want to create and edit images without sending that work to a cloud service, including artists developing character concepts or trying out apparel designs.
8.9KUpdated 3 weeks agoApache-2.0
Docker#Batch processing#ControlNet#Distributed execution
BentoML is a Python framework for developers turning AI models into services on their own hardware or servers. It supports self-hosted inference APIs and multi-model applications, with Apache 2.0 licensing. You can develop and debug locally, then deploy the services in Docker containers, on Kubernetes, or in your own cloud.
10.1KUpdated 2 weeks agoApache-2.0
Docker#Distributed execution#Hugging Face integration#LoRA
OpenRLHF is a self-hosted Python framework for researchers and teams training language models with human feedback or custom rewards. It runs on your own NVIDIA GPU hardware, with Docker support and distributed training across servers. It's open source under Apache 2.0.
33.9KUpdated 2 weeks agoAGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#Multilingual
SillyTavern is a locally installed LLM frontend for AI hobbyists who want detailed control over character chats and prompts. It builds on TavernAI as an independently developed fork and brings text models, image generation and voice into one interface. It's open source under AGPL-3.0.
12.4KUpdated 23 hours ago
iOS · Android · Docker · Web#Batch processing
Inbox Zero is an AI email assistant you can self-host with Docker or use as a hosted service. It's for people handling busy Gmail, Google Workspace, or Microsoft Outlook inboxes who want help sorting mail and preparing replies. It works alongside your existing email client.
7.7KUpdated 5 days agoMIT
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Hugging Face integration
mistral.rs is an open source inference engine for running models on your own computer or self-hosted server. It's for developers building AI applications and people who want local chat, multimodal models and agent tools in the same runtime. The Rust project uses the MIT license.