12.3KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval
ShellGPT is an AI terminal assistant for developers and people who work with shell commands. It runs on Linux, macOS and Windows, turning plain-language requests into commands suited to your operating system and shell. You can review, explain or execute its suggestions, and its Bash and Zsh integrations put generated commands into the terminal input line for editing.
784Updated 4 months agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multi-user access#Multimodal input#Persistent memory
Agnai is a self-hosted AI roleplay chat app for people who want to create fictional characters and talk with them alone or in a group. A conversation can include multiple people and multiple bots. It builds on early work from Galatea-UI by PygmalionAI and uses the AGPL-3.0 open-source license.
1.7KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#Multilingual#Persistent memory
RisuAI is an open-source AI roleplay client for people who want to create characters, build fictional worlds and chat with several characters together. It runs on Windows, macOS, Linux, Android and in a browser. You can also host the web app yourself with Docker.
1.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux#Ollama integration#RAG#Works offline
tlm is an open source terminal assistant for people who want help writing shell commands and understanding unfamiliar ones. It runs models on your workstation through Ollama, so command assistance doesn't depend on a cloud service. It works offline and requires no API key or subscription.
4.5KUpdated 7 months agoMIT
macOS · Windows · Linux#MCP#Ollama integration#OpenAI-compatible API
mods is a command-line AI tool for people who want to ask questions about command output or use model responses in shell pipelines. It connects to local LLMs through LocalAI as well as cloud services. The project is archived and no longer maintained.
desktop.backyard.aiDesktop Chat Apps
macOS · Windows#llama.cpp backend
Backyard AI's desktop app uses llama.cpp, a runtime for running language models locally. It's for people looking for a desktop AI app on Windows or macOS, with separate Mac downloads for Intel and Apple Silicon hardware.
5.4KUpdated 3 months ago
macOS · Windows · Linux#MCP#Multilingual#Ollama integration
5ire is a free desktop AI assistant that combines chat, a local document knowledge base and MCP tools. It runs on macOS, Windows and Linux, with Mac downloads for Apple Silicon and Intel. It's for people who want to use their own documents and external tools alongside conversations with local or cloud models.
13KUpdated 11 months agoApache-2.0
Windows · Web#Hugging Face integration#LoRA#Multimodal input
CogVideoX is a family of downloadable video generation models for developers, researchers and creators who want to generate clips on their own hardware. It turns English text prompts into video, animates a supplied image and can continue an existing video. A local Gradio web interface provides a browser front end for generation.
5.7KUpdated 1 year agoApache-2.0
Windows · Docker · Web#llama.cpp backend
Serge is a self-hosted chat interface for people who want to run language models on their own hardware and talk to them in a browser. The project is archived and no longer maintained. It uses llama.cpp to run models locally, with Alpaca as a named chat model and LLaMA also referenced in its memory requirements.
4.8KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
Lollms WebUI is a local, single-user AI interface for people who want text chat and media generation in one place. It runs on Windows, macOS and Linux, with Docker support, and lets writers, developers and other users choose models and task-specific personalities. It's free and open source under Apache 2.0. The project receives minimal maintenance.
33.3KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Ollama integration
Chatbot UI is a browser-based AI chat app for people who want to host their own interface and choose between local models and cloud providers. It connects to Ollama for models running on your hardware, alongside services such as OpenAI, Azure OpenAI, Anthropic and Google Gemini. The app is open source under the MIT license.
1.9KUpdated 3 weeks agoAGPL-3.0
macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration
Sonar is a self-hosted inference engine for developers and teams serving Hugging Face-compatible language and multimodal models on their own hardware. Based on vLLM, it adds model and quantization formats, sampling methods, and deployment features. It's open source under AGPL-3.0.
qualcomm/GenieXInference Libraries and Bindings
macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend
Nexa SDK is an on-device AI inference framework for developers building applications that process text, images or audio on users' hardware. It runs models locally across CPUs, GPUs and NPUs, with a shared interface for different backends. Its scope includes language and vision models, speech recognition, speech synthesis and image generation.
16.2KUpdated 1 day agoApache-2.0
Windows · iOS · Android#Image-to-image#Multimodal input#ONNX
MNN is a lightweight C++ framework for developers who want AI models to run on phones, PCs and embedded devices. It handles inference and training on the device, with a focus on small application footprints and hardware acceleration. The project is open source under Apache 2.0, and Alibaba uses it in apps including Taobao, Youku and DingTalk.
3.1KUpdated 2 days agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Multi-agent workflows
BotSharp is a self-hosted framework for .NET developers building AI agents into business applications. Written in C#, it runs on Windows, Linux and macOS and is open source software under Apache 2.0. Its plugin design lets teams choose their model provider, storage and interface while keeping agent coordination in the same framework.
8.1KUpdated 1 day agoMIT
Windows · Linux · iOS · Android · Docker · Web#Multi-user access#ONNX#Semantic search
LibrePhotos is a self-hosted photo manager for people who want to organize a photo library on their own server, with AI tools for finding images and grouping faces. It scans files already on your storage and supports RAW photos as well as videos. The project is open source under the MIT license.
4.6KUpdated 20 hours agoMIT
macOS · Windows · Linux · Docker · Web#Visual workflows
SwarmUI is a self-hosted AI image generation interface that combines a straightforward generation screen with direct access to ComfyUI's node-based workflows. It's for people who want to run creative models on their own hardware, with room to build more detailed workflows as their needs grow.
15.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux · Docker#Multilingual
Unstructured is a local document processing library for developers building LLM applications and document ingestion pipelines. It turns PDFs, Word documents, HTML, emails and images into document elements that applications can use. The Python library is open source under Apache 2.0 and runs on your own hardware, including through Docker images for x86_64 and Apple Silicon.
2.1KUpdated 4 months agoGPL-3.0
Windows
Flowframes is a Windows desktop app that uses local AI to generate intermediate frames for smoother video and animation. It's for people working with camera footage, 2D animation or rendered video who want a graphical interface for interpolation. The app is open source under GPL-3.0 and available as donationware, with free builds.
4.3KUpdated 4 weeks agoApache-2.0
macOS · Windows · Linux · iOS · Android#Batch processing#Semantic search
USearch is an open-source similarity search library for developers building semantic search, recommendation systems, and other applications that compare vectors. It runs within your application on your own hardware or server. Its compact C++ core supports custom definitions of similarity, including comparisons between combined image and text embeddings or geospatial data.
606Updated 4 days agoApache-2.0
Windows#ControlNet#GGUF#Inpainting
Amuse combines AI generation with media editing in a Windows app that runs models on your own hardware. It's for people who want to create images, video, audio and text locally, then work on the results in the same application. Its editor also accepts existing local video.
2.3KUpdated 2 years ago
Windows · Web#Hugging Face integration
Deforum is a free Stable Diffusion tool for artists and developers who want to generate images and animations on their own hardware. The project is no longer maintained. Its code uses MIT, while downloaded Stable Diffusion models retain their own weight licenses. Its Python script and Jupyter notebook support local generation, while Google Colab and Replicate offer cloud execution.
15.7KUpdated 12 months agoMIT
Windows · Docker#Code execution#Git integration#Human approval
Plandex is an open source, terminal-based AI coding agent for developers working on changes that span many files. Its review sandbox holds generated edits apart from your project files so you can inspect the accumulated changes before applying them. You can run its server locally through Docker or host it on your own server. The code uses the MIT license.
88.8KUpdated 2 months agoMIT
macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API
NextChat is a self-hosted AI chat interface for people who want one place to use their own LLM server and cloud models. The web and desktop project is open source under the MIT license. You can host it with Docker or on Vercel, and desktop clients run on macOS, Windows and Linux.
1.3KUpdated 7 months agoApache-2.0
Windows · Docker#Hugging Face integration#Multimodal input#OpenAI-compatible API
JoyCaption is an open-weight image captioning model for people preparing datasets to train or fine-tune diffusion models. It runs on your own GPU and covers both SFW and NSFW images, including photography, anime, digital art and furry artwork. Automated captions reduce the need to write descriptions by hand or find images that already have usable text.
11KUpdated 1 year agoApache-2.0
macOS · Windows · Linux · Web#Hugging Face integration#Multilingual#Voice cloning
Spark-TTS is a local text-to-speech system that can copy a voice from reference audio or create a synthetic speaker with adjustable vocal traits. It's for developers and researchers building speech applications, including personalized narration, assistive technology, and language research. The Python and PyTorch code is open source under Apache 2.0.
23.3KUpdated 1 year agoApache-2.0
macOS · Windows · Web#Batch processing#Image-to-image#Inpainting
IOPaint is a free, self-hosted AI image editor for people who want to remove objects, replace parts of a picture or extend it beyond its original edges on their own hardware. The project is archived and no longer maintained. It's open source under Apache 2.0, with a browser interface and support for CPU, GPU and Apple Silicon hardware.
14.9KUpdated 2 years agoApache-2.0
macOS · Windows · Docker#Streaming inference#Voice cloning
Tortoise TTS is a local text-to-speech system for developers and creators who want speech with varied voices and natural pacing. It uses reference audio clips to guide a custom voice, with an emphasis on expressive rhythm and intonation.
8.9KUpdated 8 months agoApache-2.0
Windows · Linux · Docker#Distributed execution#GGUF#Hugging Face integration
Intel IPEX-LLM is a library for developers running or fine-tuning models on Intel hardware. The project is archived and no longer maintained. Intel reports known security issues and no longer accepts patches or provides updates. The code is open source under Apache 2.0.
4.6KUpdated 7 months agoMIT
Windows · Linux#Batch processing#Quantization#Speculative decoding
ExLlamaV2 is a local LLM inference library for developers and people hosting models on their own consumer GPUs. ExLlamaV2 is archived and no longer maintained; development continues in ExLlamaV3. The V2 library is free and open source under the MIT license, runs on Windows and Linux, and uses NVIDIA GPUs through CUDA. It supports multiple GPUs.
2KUpdated 7 months agoMIT
macOS · Windows · Linux · Web#MCP#RAG
NotebookLlama is a self-hosted NotebookLM alternative for people who want to work with documents through an app they can host and modify. Its browser interface runs locally, while LlamaCloud provides document extraction and indexing. It requires cloud services.
3.5KUpdated 20 hours agoApache-2.0
macOS · Windows · Linux · iOS · Android · Web#Agent Skills#Hugging Face integration#Multimodal input
LiteRT is Google's open-source framework for developers building AI into apps that run on users' own devices. It succeeds TensorFlow Lite and covers model conversion, optimization and local inference. It's licensed under Apache 2.0.
23.9KUpdated 6 days ago
macOS · Windows · Linux · iOS · Android · Web#ONNX#Quantization
ncnn is a C++ framework for developers building on-device AI into mobile, desktop and embedded applications. Its focus is running neural networks with a small memory footprint and no third-party runtime dependencies. Models run on the target device's CPU or a supported Vulkan GPU.
5.7KUpdated 2 weeks agoMIT
macOS · Windows · Linux · Docker#Home Assistant integration#MCP#Multi-agent workflows
GLaDOS is a local AI voice assistant modeled on the sarcastic character from Valve's Portal games. It's for people who want a conversational companion on their own hardware, with camera awareness and connections to home automation. The Python project is open source under the MIT license and runs on Linux and Windows. macOS support is experimental.
1.2KUpdated 12 months agoMIT
macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Hollama is an open-source LLM chat app whose interface runs entirely in your browser. It's for people who want to chat with local AI through Ollama or connect to OpenAI servers, with support for multiple server connections. The interface stores data locally in the browser; the connected server handles model requests, so where inference runs depends on the server you choose.
2.8KUpdated 9 months agoApache-2.0
Windows · Linux#Batch processing#ONNX#Voice activity detection
openWakeWord is a Python library for developers building voice interfaces that listen locally for a chosen word or phrase. It includes English models for triggers such as "hey jarvis" and "alexa", plus phrases for weather and timers. The code uses Apache 2.0. Included pretrained models use CC-BY-NC-SA-4.0, which restricts commercial use.