1.2KUpdated 12 months agoMIT
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Hollama is an open-source LLM chat app whose interface runs entirely in your browser. It's for people who want to chat with local AI through Ollama or connect to OpenAI servers, with support for multiple server connections. The interface stores data locally in the browser; the connected server handles model requests, so where inference runs depends on the server you choose.
4.4KUpdated 20 hours agoMIT
macOS · Windows · Linux · Web · JetBrains#Agent Client Protocol#Code execution#Git integration
gptme is a self-hosted AI agent that works directly in your terminal, with access to your files and installed tools. It's for developers who want a coding assistant in their own environment, and people who want an agent for data analysis or other knowledge work. The software is free under the MIT license and doesn't require a gptme account.
7.5KUpdated 2 days agoApache-2.0
#Distributed execution#Hugging Face integration#OpenAI-compatible API
OpenCompass is an open-source LLM evaluation platform for researchers, model developers and teams comparing models for their applications. Its Python framework evaluates local models and cloud APIs within the same experiment, so teams can compare candidates on shared benchmarks. It uses the Apache 2.0 license.
5.1KUpdated 22 hours ago
macOS · Windows · Linux · iOS · Android · Web#MLX#Multimodal input#OpenAI-compatible API
ExecuTorch is PyTorch's runtime for developers building AI into mobile apps, desktop software and embedded devices. It runs models on the user's hardware, with support for Android, iOS, Linux, macOS and Windows, as well as microcontrollers. Developers can reuse a PyTorch model across targets, though hardware-specific deployments need their own exported model files.
8.2KUpdated 3 days agoMIT
Web · Browser Extension#LM Studio integration#Ollama integration#OpenAI-compatible API
Page Assist brings local AI chat into your browser, with a sidebar for conversations alongside a webpage and a separate tab for a ChatGPT-style interface. It's for people who already run models on their own hardware and want to ask questions about what they're reading without leaving the page.
11KUpdated 1 day agoApache-2.0
Docker · Web#llama.cpp backend#MCP#Multi-user access
HuggingChat UI is the open-source chat application behind Hugging Face's hosted HuggingChat. You can run it on your own computer or server and connect it to a local LLM backend or a cloud provider. It's for people and teams who want a browser-based ChatGPT alternative with control over the chat service and where its data lives. The code uses the Apache 2.0 license.
752Updated 3 weeks agoApache-2.0
Linux · Docker · Web · Browser Extension#Batch processing#OpenAI-compatible API#Quantization
Meme Search is a free, self-hosted web app for people who want to find memes in their own collection by image content or text. It runs locally through Docker and uses AI descriptions to make images searchable, even when their filenames aren't useful. The code uses the Apache 2.0 license.
1.5KUpdated 3 weeks agoMIT
Docker · Web#Ollama integration#OpenAI-compatible API#Streaming inference
Unmute adds spoken conversation to text LLMs using Kyutai's speech recognition and speech synthesis models. It's for developers who want a self-hosted voice interface while keeping their choice of language model. The project uses the MIT license, and a hosted browser demo is available at Unmute.sh.
2.6KUpdated 21 hours agoApache-2.0
Web#OpenAI-compatible API#Prompt caching
vLLM Production Stack is an open source inference stack for teams serving LLMs on their own Kubernetes GPU clusters. It brings request routing and monitoring around vLLM, so applications can move from one serving instance to a distributed deployment without changing their code. It requires a GPU-enabled Kubernetes environment.
9.4KUpdated 3 weeks agoMIT
Docker#Batch processing#GGUF#Hugging Face integration
SenseVoice is a local speech recognition model that adds language, emotion and sound-event tags to transcriptions. It's for developers building voice applications or analyzing recordings on their own hardware, particularly those working with Mandarin and Cantonese. The project is open source under the MIT license.
goodsnooze.gumroad.comDictation and Voice Typing
macOS · iOS#Batch processing#Multilingual#Ollama integration
MacWhisper is a native macOS transcription app for people working with interviews, lectures, meetings and other recorded audio. It runs speech recognition on your own Mac, so local transcription keeps audio on your device. It also offers cloud transcription through services such as OpenAI, ElevenLabs and Deepgram, which send audio off your machine.
1.7KUpdated 2 days agoApache-2.0
#Batch processing#Code execution#Multimodal input
Curator is a Python library for developers preparing LLM training datasets or extracting structured records from existing data. It supports local inference through Ollama and vLLM alongside cloud model APIs, so the same data pipeline can use models on your hardware or a hosted provider. It's open source under Apache 2.0.
12.5KUpdated 4 months agoApache-2.0
Docker · Web#Hugging Face integration#OpenAI-compatible API
OpenLLM is a self-hosted LLM server for developers who want to connect their applications to models running on their own hardware or servers. Its OpenAI-compatible API works with clients built for that interface, including the OpenAI Python client and LlamaIndex. The project is open source under the Apache License 2.0.
18.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#MLX#Multilingual
VideoLingo is a self-hosted video translation app for creators and educators who need bilingual subtitles or dubbed versions of their videos. It brings transcription, translation and subtitle timing into one browser interface, with dubbing as an optional output. The project is open source under Apache 2.0; a separate hosted service offers subtitle translation and dubbing.
3.2KUpdated 5 days agoApache-2.0
macOS · Linux · Docker#GGUF#llama.cpp backend#MCP
Harbor is a CLI and companion app for people experimenting with AI on their own hardware. It manages a local LLM development environment, connecting model backends to chat interfaces and supporting services so you don't have to configure each connection yourself. It's open source under Apache 2.0.
1.1KUpdated 2 years agoGPL-3.0
macOS · Windows · Linux#Multilingual#OpenAI-compatible API#Quantization
WhisperWriter turns microphone speech into text and types it into the window you're working in. It's for people who want voice input in their existing desktop apps, with a choice between transcription on their own computer and an external service. The Python app runs on Windows, macOS and Linux and uses the GPL-3.0 open-source license.
1.5KUpdated 2 weeks agoApache-2.0
#Home Assistant integration#Multimodal input#Ollama integration
LLM Vision is a free, open-source Home Assistant integration for people who want their smart home to interpret what cameras see. It uses multimodal LLMs to describe images, video files, live feeds and Frigate events, then uses those results in notifications and automations.
37.5KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#LLM tracing#MCP#Multimodal input
Claude Code Router is an open-source local model gateway for developers who use coding agents and want to manage their model providers in one place. It runs on macOS, Windows and Linux, with Docker and a CLI with a browser interface also available. The project uses the MIT license.
3.1KUpdated 1 week agoMIT
macOS · Linux · Docker · Web#Ollama integration#OpenAI-compatible API#Speaker diarization
Scriberr is a free, open source transcription app for people who want to keep meeting recordings and voice notes on their own hardware. It turns audio and video into text locally, with offline transcription once the models are downloaded. The MIT license allows you to use and modify it.
2KUpdated 6 months agoAGPL-3.0
macOS · Windows · Linux#Code execution#LM Studio integration#MCP
Witsy is an AGPL-3.0 desktop AI assistant for macOS, Windows and Linux that connects MCP tools to local and cloud models. It's for people who want document chat, writing help and voice features in one app, with a choice of where their models run.
1.8KUpdated 20 hours agoApache-2.0
Linux · Docker#Distributed execution#Hugging Face integration#Multilingual
NeMo Curator is an open source Python toolkit for ML engineers and data teams preparing AI training datasets on their own hardware. It handles text, images, video and audio, with reusable pipelines that can run on a laptop or scale across a multi-node Ray cluster. NVIDIA uses it to prepare data for Nemotron models.
18KUpdated 1 day agoApache-2.0
Web#Code execution#LM Studio integration#MCP
LangBot is a self-hosted AI agent platform for teams that want bots in the messaging apps their customers, coworkers or communities already use. It connects Slack, Discord, Telegram, WeChat and other chat services to models and AI workflows, with a browser dashboard for managing bots across platforms.
3.3KUpdated 3 weeks agoMIT
Windows · Docker · Web#OpenAI-compatible API
TTS WebUI brings local text-to-speech, music generation and audio processing into one browser interface. It's for people creating spoken audio or music, and for developers who want to add speech to a self-hosted chat app. The interface combines Gradio and React, with extensions that let you choose which audio models to use.
1.4KUpdated 3 days ago
#GGUF#Home Assistant integration#Hugging Face integration
Home-LLM connects Home Assistant to language models running on your own hardware, so you can control smart devices through voice or chat. It's for Home Assistant users who want natural language control without relying on a cloud service or subscription. The project pairs a custom integration with small models trained specifically for smart home commands.
1.4KUpdated 1 day ago
Docker · Web#Git integration#Multi-user access#OpenAI-compatible API
Kodus is an AI code review tool for engineering teams that want automated pull request feedback while choosing where the reviewer runs and which model it uses. Its open source core uses the AGPL license and can be self-hosted with Docker Compose or on Kubernetes. Kodus also offers a hosted cloud service that manages the infrastructure.
836Updated 2 months agoMIT
Docker#LM Studio integration#Multi-user access#Multimodal input
llmcord is a self-hosted Discord bot for people who want to share LLM conversations with friends or a community. You can run the Python bot on your own machine or server, including through Docker, and connect it to local models or cloud providers. It's open source under the MIT license.
181Updated 7 months ago
macOS · Windows · Linux#Multilingual#Ollama integration#OpenAI-compatible API
LocalWriter brings local LLM writing assistance into LibreOffice Writer for people who want to draft and revise text inside their documents. It runs on macOS, Windows and Linux and connects to a separate model runner, including Ollama and text-generation-webui. With a backend on your own machine, text processing stays local.
295Updated 2 days agoApache-2.0
Linux · Docker#OpenAI-compatible API#Wake word detection#Works offline
OpenVoiceOS is a free, open-source voice AI platform for people building their own smart speakers or adding voice control to devices. Its core builds on a fork of MycroftAI/mycroft-core, and most classic Mycroft skills also work with it. The project uses the Apache 2.0 license, which permits personal and commercial use.
68.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux#Agent Client Protocol#Code execution#Human approval
Open Interpreter is a terminal coding assistant built around open-weight models, with model-specific behavior for Kimi, Qwen and DeepSeek. It's a fork of OpenAI's Codex for developers who want to choose their model provider while keeping a familiar agent interface. The project uses the Apache 2.0 license.
70.7KUpdated 8 months agoMIT
Windows · Docker#Code execution#Multi-agent workflows#Ollama integration
MetaGPT organizes AI agents into a software development team, with roles for product management, architecture, project management and engineering. It's a self-hosted Python framework for developers who want to generate software projects from plain-language requirements or build their own collaborative agents.
2.2KUpdated 6 days agoApache-2.0
#Ollama integration#OpenAI-compatible API#Tool calling
any-llm is a Python library for developers who want the same application to work with local LLM servers and cloud providers. It connects to Ollama and custom OpenAI-compatible endpoints, alongside OpenAI, Anthropic, Mistral and Azure / Microsoft Foundry. A shared interface reduces the provider-specific code needed to try another model or change where inference runs.
677Updated 5 months agoMIT
macOS#Multimodal input#Ollama integration#OpenAI-compatible API
Obsidian Local GPT brings AI writing assistance into Obsidian, with local Ollama models for private, offline use or connections to OpenAI-compatible services. It's for people who want help with their notes while keeping control over where the AI runs. The plugin is open source under the MIT license.
49.1KUpdated 2 days agoAGPL-3.0
Linux · Docker · Web#Multi-user access#Multimodal input#OpenAI-compatible API
New API is a self-hosted AI gateway for developers and teams that want several model providers behind one service. It builds on One API and converts between OpenAI Chat Completions, Responses, Anthropic Messages and Gemini formats, so apps and agents can switch providers without changing each client's connection settings.
1.7KUpdated 2 days agoMIT
#LM Studio integration#Ollama integration#OpenAI-compatible API
LLPhant is an MIT-licensed PHP framework for adding language models, embeddings and vector databases to Symfony and Laravel applications. It requires PHP 8.1 or later and is installed through Composer.
13.8KUpdated 1 month agoApache-2.0
Browser Extension#Multi-agent workflows#Ollama integration#OpenAI-compatible API
Nanobrowser is an open-source AI agent that automates web tasks inside Chrome or Edge. It's for people who want to delegate repetitive browsing or research while choosing which models handle the work. The extension runs in your browser and is an alternative to OpenAI Operator, with an Apache 2.0 license.
38.7KUpdated 11 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Langchain-Chatchat is a self-hosted application for asking questions about your own documents and using AI agents. It focuses on Chinese-language use and open models, with a fully offline setup that can keep documents and model processing on your hardware. Its code is open source under Apache 2.0.