4.4KUpdated 20 hours agoMIT
macOS · Windows · Linux · Web · JetBrains#Agent Client Protocol#Code execution#Git integration
gptme is a self-hosted AI agent that works directly in your terminal, with access to your files and installed tools. It's for developers who want a coding assistant in their own environment, and people who want an agent for data analysis or other knowledge work. The software is free under the MIT license and doesn't require a gptme account.
5.1KUpdated 21 hours ago
macOS · Windows · Linux · iOS · Android · Web#MLX#Multimodal input#OpenAI-compatible API
ExecuTorch is PyTorch's runtime for developers building AI into mobile apps, desktop software and embedded devices. It runs models on the user's hardware, with support for Android, iOS, Linux, macOS and Windows, as well as microcontrollers. Developers can reuse a PyTorch model across targets, though hardware-specific deployments need their own exported model files.
2.1KUpdated 3 days ago
Windows · Linux#LoRA#Quantization
Musubi Tuner is a Python toolkit for training LoRA adapters for image and video generation models on your own hardware. It's aimed at people who want to customize these models using their own datasets and are comfortable working with training scripts. It also includes image and video generation scripts for supported architectures.
11.1KUpdated 1 month agoGPL-3.0
macOS · Windows · Linux · Android#RAG#Semantic search
Blinko is a self-hosted AI note app for quickly capturing thoughts and retrieving them later. Notes and data are stored in your own environment, with plain-text storage and Markdown formatting. The project is licensed under GPL-3.0.
2KUpdated 2 days agoGPL-3.0
Windows · Linux#Distributed execution#LoRA#Quantization
diffusion-pipe is a local diffusion model training tool for people fine-tuning image and video models on their own GPU hardware. Its main distinction is that it can divide a model across several GPUs when it won't fit on one, while also distributing training work across GPUs. The Python project is open source under GPL-3.0 and uses DeepSpeed.
1.4KUpdated 12 months agoGPL-3.0
macOS · Windows · Linux#Batch processing#Multimodal input
TagGUI is a desktop app for people preparing image datasets for generative AI training. It combines local AI captioning with manual tag editing, so you can generate descriptions and correct them in the same workspace. It's open source under GPL-3.0 and runs on Windows, Linux and macOS, though macOS doesn't have a packaged release.
9.4KUpdated 2 days agoMIT
macOS · Windows · Linux · Android · Docker · Web#Batch processing#LM Studio integration#MCP
xberg, formerly Kreuzberg, is a local document extraction engine for developers building AI search, document processing, and retrieval-augmented generation applications. It reads PDFs, Office files, scanned images, email, and nested archives, extracting text, tables, images, and metadata through one shared engine. It's open source under MIT.
1.3KUpdated 1 day ago
macOS · Windows · Linux#GGUF#Hugging Face integration#LoRA
GPTQModel is a Python toolkit for developers compressing LLMs and running them on their own hardware or servers. It brings model calibration, compression, quality checks and inference into one API, so teams can compare quantization methods without adopting a separate tool for each one.
11.2KUpdated 1 month ago
macOS · Windows · Linux · iOS · Android · Web#Multilingual#Streaming inference
Moonshine is an on-device AI toolkit for developers building voice agents and applications that listen and speak. It combines speech to text, intent recognition and text to speech in one library. Voice processing stays on the device, and you don't need an account or API keys.
7.2KUpdated 6 days agoApache-2.0
Windows · Linux#ControlNet#Inpainting#LoRA
sd-scripts is a collection of Python scripts for training and generating images with models on your own hardware. It's aimed at people who want to customize image models through LoRA training or deeper fine-tuning and are comfortable working with scripts. The project is open source under the Apache 2.0 license.
2.8KUpdated 1 day agoApache-2.0
Windows · Linux · Docker · Web#LLM tracing#Ollama integration#Prompt versioning
OpenLIT is a self-hosted platform for developers who need to understand how their LLM applications and AI agents behave. It connects model calls with tool activity, retrieval and agent steps, so teams can investigate errors and compare cost, latency and output quality across a workflow.
4.7KUpdated 1 week agoApache-2.0
macOS · Windows · Linux · Docker#Hybrid search#Reranking#Semantic search
Infinity is a self-hosted database for developers building search and retrieval-augmented generation (RAG) into LLM applications. It combines embedding search with full-text search and structured filters, so an application can retrieve relevant records through both meaning and exact terms.
8.3KUpdated 3 years agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multi-user access#Role-based access
CompreFace is a self-hosted face recognition service for developers who want to add facial identification to an application without building or training their own machine learning system. It runs as a Docker-based server on your hardware or in a cloud deployment you manage. It's free and open source under the Apache 2.0 license.
23.8KUpdated 19 hours agoMPL-2.0
macOS · Windows · Linux · iOS · Android#Multilingual#Persistent memory#RAG
Brave Leo is an AI assistant built into the Brave browser, with Bring Your Own Model support for people who want to use their own local or remote models while browsing. It can work with third-party APIs as well as Brave's hosted model choices. The browser runs on macOS, Windows, Linux, Android, and iOS.
30KUpdated 10 months agoApache-2.0
Windows#Multilingual
EasyOCR is a Python OCR library for developers who want to extract text from images on their own hardware. It reads text in photographs and dense documents, so it can serve both scene-text recognition and document processing. It's open source under Apache 2.0.
18.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#MLX#Multilingual
VideoLingo is a self-hosted video translation app for creators and educators who need bilingual subtitles or dubbed versions of their videos. It brings transcription, translation and subtitle timing into one browser interface, with dubbing as an optional output. The project is open source under Apache 2.0; a separate hosted service offers subtitle translation and dubbing.
38.6KUpdated 2 months agoMIT
Windows · Linux · Web#Hugging Face integration#ONNX#Voice conversion
RVC WebUI is a local AI voice conversion tool for people who want to train a custom voice, change the voice in a recording, or use a live voice changer. It runs on Windows and Linux, including Ubuntu servers, with a browser interface for training and conversion and a separate interface for live use. It's free and open source under the MIT license.
2.4KUpdated 2 years agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Hugging Face integration
AllTalk TTS generates speech on your own computer. The project recommends v2 for most users; the saved documentation below describes v1, built on Coqui TTS and XTTSv2 models. It's for people adding voices to AI conversations or producing spoken audio from longer texts. It runs as a standalone application or alongside Text-generation-webui, with support for Windows, Linux and macOS.
1.1KUpdated 2 years agoGPL-3.0
macOS · Windows · Linux#Multilingual#OpenAI-compatible API#Quantization
WhisperWriter turns microphone speech into text and types it into the window you're working in. It's for people who want voice input in their existing desktop apps, with a choice between transcription on their own computer and an external service. The Python app runs on Windows, macOS and Linux and uses the GPL-3.0 open-source license.
37.5KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#LLM tracing#MCP#Multimodal input
Claude Code Router is an open-source local model gateway for developers who use coding agents and want to manage their model providers in one place. It runs on macOS, Windows and Linux, with Docker and a CLI with a browser interface also available. The project uses the MIT license.
2KUpdated 6 months agoAGPL-3.0
macOS · Windows · Linux#Code execution#LM Studio integration#MCP
Witsy is an AGPL-3.0 desktop AI assistant for macOS, Windows and Linux that connects MCP tools to local and cloud models. It's for people who want document chat, writing help and voice features in one app, with a choice of where their models run.
1.9KUpdated 3 months agoMIT
macOS · Windows · Linux#Distributed execution#llama.cpp backend#Quantization
Augmentoolkit turns your documents into training data for a custom LLM that learns a particular subject. It's for researchers, developers and hobbyists who want models trained on their own material, such as research papers or fictional lore. The Python toolkit is open source under the MIT license and runs on macOS and Linux, with WSL recommended for Windows.
28.3KUpdated 1 day ago
macOS · Windows · Linux · Docker · Web#Code execution#MCP
Chat2DB Community is a database client and AI SQL workspace for developers, database administrators and analysts. It runs on Windows, macOS and Linux, with Docker and browser access also available. Its AI assistant connects to a model you configure to generate, explain and optimize SQL. Where AI requests are processed depends on that model connection.
34.4KUpdated 2 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#MCP
ChatDev is a self-hosted platform for building teams of AI agents through a visual workflow editor. It's for people who want agents to collaborate on research, data analysis or software projects without writing the orchestration code themselves. You define each agent's role and how information passes between them.
3.3KUpdated 3 weeks agoMIT
Windows · Docker · Web#OpenAI-compatible API
TTS WebUI brings local text-to-speech, music generation and audio processing into one browser interface. It's for people creating spoken audio or music, and for developers who want to add speech to a self-hosted chat app. The interface combines Gradio and React, with extensions that let you choose which audio models to use.
1.2KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker · Web#Semantic search
HomeGallery is a self-hosted web gallery for people who want to browse their personal photos and videos while keeping their existing files and folders. It brings multiple media directories into one gallery, with a mobile-friendly browser interface and PWA support. It's open source under MIT.
2.3KUpdated 1 day agoMPL-2.0
macOS · Windows · Linux · Docker#Agent Skills#Batch processing#Multi-user access
dstack is a self-hosted orchestration tool for AI teams managing compute across GPU clouds and their own servers. It puts cluster management, training jobs and model inference behind one interface, so teams can use different providers and accelerators without maintaining a separate workflow for each environment. It's open source under the Mozilla Public License 2.0.
1KUpdated 7 days agoMIT
macOS · Windows · Linux#Batch processing#Multimodal input#Semantic search
rclip searches image folders by their visual content, so you can find photos without adding tags or importing them into a photo library. It's a local AI tool for people who keep image collections on their own computers or servers and prefer working in the terminal. It runs on Linux, Windows and Apple Silicon macOS, and it's open source under the MIT license.
16.2KUpdated 19 hours agoGPL-3.0
macOS · Windows · Linux#Multimodal input
LabelMe is a desktop image annotation app for people preparing computer vision datasets. It combines manual drawing with AI assistance for outlining objects and creating labels from text. It runs on 64-bit macOS, Windows and Linux..
181Updated 7 months ago
macOS · Windows · Linux#Multilingual#Ollama integration#OpenAI-compatible API
LocalWriter brings local LLM writing assistance into LibreOffice Writer for people who want to draft and revise text inside their documents. It runs on macOS, Windows and Linux and connects to a separate model runner, including Ollama and text-generation-webui. With a backend on your own machine, text processing stays local.
9.8KUpdated 5 months agoMIT
macOS · Windows · Linux#Batch processing#GGUF#Hugging Face integration
PowerInfer is a local LLM inference engine for developers and researchers who want to run large models on a PC with a consumer GPU. It splits work between the CPU and GPU to reduce GPU memory demands and data transfers. The code is open source under the MIT license.
62.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection
GPT-SoVITS is a local text-to-speech and voice cloning tool. It can generate speech from a short reference recording or fine-tune a model for a custom voice. The source code uses the MIT license.
3.4KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker#Code execution#MCP#Tool calling
Postgres MCP Pro is a self-hosted MCP server that gives AI assistants access to PostgreSQL alongside tools for diagnosing slow queries and testing index recommendations. It's for developers and database maintainers who want their assistant to work with the database's actual schema, query statistics and execution plans. It runs through Docker or Python and is open source under the MIT license.
541Updated 7 months agoGPL-3.0
macOS · Windows · Linux · iOS · Android#Multimodal input#Ollama integration
Reins is an open-source chat app for people using self-hosted LLMs through Ollama. It runs on iOS, Android, macOS, Linux and Windows, giving Ollama users a mobile and desktop interface for experimenting with models. Reins is the client; Ollama provides the model backend.
21.1KUpdated 4 days ago
macOS · Windows · Linux · Docker#ONNX#Voice conversion
Voice Changer (w-okada), also called VCClient, converts your voice as you speak using AI voice models. It's for people who want live voice conversion on their own computer, including those recording gaming commentary while running demanding software. Processing can stay local.
14.1KUpdated 3 years ago
macOS · Windows · Linux · Web#Multimodal input#Works offline
SadTalker generates talking head videos from a single portrait and an audio recording. The project states an Apache 2.0 license and removal of its earlier noncommercial restriction. It runs locally on Windows, Linux and macOS, and suits creators who want to animate a face without recording a person on camera. Its animation includes facial expressions and head movement, with examples covering speech and singing in different languages.