11.2KUpdated 1 month ago
macOS · Windows · Linux · iOS · Android · Web#Multilingual#Streaming inference
Moonshine is an on-device AI toolkit for developers building voice agents and applications that listen and speak. It combines speech to text, intent recognition and text to speech in one library. Voice processing stays on the device, and you don't need an account or API keys.
7.2KUpdated 6 days agoApache-2.0
Windows · Linux#ControlNet#Inpainting#LoRA
sd-scripts is a collection of Python scripts for training and generating images with models on your own hardware. It's aimed at people who want to customize image models through LoRA training or deeper fine-tuning and are comfortable working with scripts. The project is open source under the Apache 2.0 license.
2.8KUpdated 1 day agoApache-2.0
Windows · Linux · Docker · Web#LLM tracing#Ollama integration#Prompt versioning
OpenLIT is a self-hosted platform for developers who need to understand how their LLM applications and AI agents behave. It connects model calls with tool activity, retrieval and agent steps, so teams can investigate errors and compare cost, latency and output quality across a workflow.
4.7KUpdated 1 week agoApache-2.0
macOS · Windows · Linux · Docker#Hybrid search#Reranking#Semantic search
Infinity is a self-hosted database for developers building search and retrieval-augmented generation (RAG) into LLM applications. It combines embedding search with full-text search and structured filters, so an application can retrieve relevant records through both meaning and exact terms.
8.3KUpdated 3 years agoApache-2.0
macOS · Windows · Linux · Docker · Web#Multi-user access#Role-based access
CompreFace is a self-hosted face recognition service for developers who want to add facial identification to an application without building or training their own machine learning system. It runs as a Docker-based server on your hardware or in a cloud deployment you manage. It's free and open source under the Apache 2.0 license.
12.3KUpdated 1 year agoMIT
Linux#Multimodal input#Structured output
Zerox is an MIT-licensed OCR library for developers preparing documents for AI applications. Its Node.js and Python packages run on your own machine or server, while cloud vision models read the document pages and produce Markdown. Document conversion happens locally, but page images go to the selected model provider, so this workflow needs internet access and provider credentials.
23.8KUpdated 19 hours agoMPL-2.0
macOS · Windows · Linux · iOS · Android#Multilingual#Persistent memory#RAG
Brave Leo is an AI assistant built into the Brave browser, with Bring Your Own Model support for people who want to use their own local or remote models while browsing. It can work with third-party APIs as well as Brave's hosted model choices. The browser runs on macOS, Windows, Linux, Android, and iOS.
18.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#MLX#Multilingual
VideoLingo is a self-hosted video translation app for creators and educators who need bilingual subtitles or dubbed versions of their videos. It brings transcription, translation and subtitle timing into one browser interface, with dubbing as an optional output. The project is open source under Apache 2.0; a separate hosted service offers subtitle translation and dubbing.
38.6KUpdated 2 months agoMIT
Windows · Linux · Web#Hugging Face integration#ONNX#Voice conversion
RVC WebUI is a local AI voice conversion tool for people who want to train a custom voice, change the voice in a recording, or use a live voice changer. It runs on Windows and Linux, including Ubuntu servers, with a browser interface for training and conversion and a separate interface for live use. It's free and open source under the MIT license.
3.2KUpdated 5 days agoApache-2.0
macOS · Linux · Docker#GGUF#llama.cpp backend#MCP
Harbor is a CLI and companion app for people experimenting with AI on their own hardware. It manages a local LLM development environment, connecting model backends to chat interfaces and supporting services so you don't have to configure each connection yourself. It's open source under Apache 2.0.
2.4KUpdated 2 years agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Hugging Face integration
AllTalk TTS generates speech on your own computer. The project recommends v2 for most users; the saved documentation below describes v1, built on Coqui TTS and XTTSv2 models. It's for people adding voices to AI conversations or producing spoken audio from longer texts. It runs as a standalone application or alongside Text-generation-webui, with support for Windows, Linux and macOS.
1.1KUpdated 2 years agoGPL-3.0
macOS · Windows · Linux#Multilingual#OpenAI-compatible API#Quantization
WhisperWriter turns microphone speech into text and types it into the window you're working in. It's for people who want voice input in their existing desktop apps, with a choice between transcription on their own computer and an external service. The Python app runs on Windows, macOS and Linux and uses the GPL-3.0 open-source license.
37.5KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#LLM tracing#MCP#Multimodal input
Claude Code Router is an open-source local model gateway for developers who use coding agents and want to manage their model providers in one place. It runs on macOS, Windows and Linux, with Docker and a CLI with a browser interface also available. The project uses the MIT license.
3.1KUpdated 1 week agoMIT
macOS · Linux · Docker · Web#Ollama integration#OpenAI-compatible API#Speaker diarization
Scriberr is a free, open source transcription app for people who want to keep meeting recordings and voice notes on their own hardware. It turns audio and video into text locally, with offline transcription once the models are downloaded. The MIT license allows you to use and modify it.
2KUpdated 6 months agoAGPL-3.0
macOS · Windows · Linux#Code execution#LM Studio integration#MCP
Witsy is an AGPL-3.0 desktop AI assistant for macOS, Windows and Linux that connects MCP tools to local and cloud models. It's for people who want document chat, writing help and voice features in one app, with a choice of where their models run.
214Updated 3 weeks agoMIT
Linux · Docker · Web#Home Assistant integration#Hugging Face integration#Multilingual
Wyoming Piper connects Piper's local text-to-speech engine to Home Assistant and other clients that use the Wyoming protocol. It's for people building a voice assistant on their own hardware who need speech generation as a self-hosted service. The project is open source under the MIT license.
1.9KUpdated 3 months agoMIT
macOS · Windows · Linux#Distributed execution#llama.cpp backend#Quantization
Augmentoolkit turns your documents into training data for a custom LLM that learns a particular subject. It's for researchers, developers and hobbyists who want models trained on their own material, such as research papers or fictional lore. The Python toolkit is open source under the MIT license and runs on macOS and Linux, with WSL recommended for Windows.
1.8KUpdated 20 hours agoApache-2.0
Linux · Docker#Distributed execution#Hugging Face integration#Multilingual
NeMo Curator is an open source Python toolkit for ML engineers and data teams preparing AI training datasets on their own hardware. It handles text, images, video and audio, with reusable pipelines that can run on a laptop or scale across a multi-node Ray cluster. NVIDIA uses it to prepare data for Nemotron models.
28.3KUpdated 1 day ago
macOS · Windows · Linux · Docker · Web#Code execution#MCP
Chat2DB Community is a database client and AI SQL workspace for developers, database administrators and analysts. It runs on Windows, macOS and Linux, with Docker and browser access also available. Its AI assistant connects to a model you configure to generate, explain and optimize SQL. Where AI requests are processed depends on that model connection.
34.4KUpdated 2 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#MCP
ChatDev is a self-hosted platform for building teams of AI agents through a visual workflow editor. It's for people who want agents to collaborate on research, data analysis or software projects without writing the orchestration code themselves. You define each agent's role and how information passes between them.
29.9KUpdated 3 weeks ago
macOS · Linux · iOS · Android · Web#Image-to-image#ONNX#Quantization
InsightFace is a face analysis toolkit for developers and teams building identity verification, access control, or face editing software. The code uses the MIT license. Its Python tools and self-hosted recognition server run inference on your own hardware. It also offers commercial models and API access for face swapping and deepfake detection.
45.1KUpdated 1 day agoAGPL-3.0
Linux · iOS · Android · Web
Logseq is an AGPL-3.0 knowledge base with an outline editor, linked references and block references. It combines note-taking with task management, PDF annotations and flashcards. Search and queries can gather related information into tables.
1.2KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker · Web#Semantic search
HomeGallery is a self-hosted web gallery for people who want to browse their personal photos and videos while keeping their existing files and folders. It brings multiple media directories into one gallery, with a mobile-friendly browser interface and PWA support. It's open source under MIT.
2.3KUpdated 1 day agoMPL-2.0
macOS · Windows · Linux · Docker#Agent Skills#Batch processing#Multi-user access
dstack is a self-hosted orchestration tool for AI teams managing compute across GPU clouds and their own servers. It puts cluster management, training jobs and model inference behind one interface, so teams can use different providers and accelerators without maintaining a separate workflow for each environment. It's open source under the Mozilla Public License 2.0.
1KUpdated 7 days agoMIT
macOS · Windows · Linux#Batch processing#Multimodal input#Semantic search
rclip searches image folders by their visual content, so you can find photos without adding tags or importing them into a photo library. It's a local AI tool for people who keep image collections on their own computers or servers and prefer working in the terminal. It runs on Linux, Windows and Apple Silicon macOS, and it's open source under the MIT license.
4.9KUpdated 3 months ago
Linux · Docker#llama.cpp backend#Multimodal input#Ollama integration
jetson-containers is a Docker container build system for developers running local AI and robotics workloads on NVIDIA Jetson hardware. It supplies prebuilt images and lets you combine AI packages into custom containers, reducing the work of assembling compatible GPU software for JetPack/L4T.
1.8KUpdated 20 hours agoMIT
macOS · Linux#Ollama integration
gollama is a terminal app for people managing a local LLM collection in Ollama on macOS or Linux. It puts model details and management actions in one keyboard-driven view, useful when you need to compare downloaded models or clear out ones you no longer use. It's open source under the MIT license.
16.2KUpdated 19 hours agoGPL-3.0
macOS · Windows · Linux#Multimodal input
LabelMe is a desktop image annotation app for people preparing computer vision datasets. It combines manual drawing with AI assistance for outlining objects and creating labels from text. It runs on 64-bit macOS, Windows and Linux..
181Updated 7 months ago
macOS · Windows · Linux#Multilingual#Ollama integration#OpenAI-compatible API
LocalWriter brings local LLM writing assistance into LibreOffice Writer for people who want to draft and revise text inside their documents. It runs on macOS, Windows and Linux and connects to a separate model runner, including Ollama and text-generation-webui. With a backend on your own machine, text processing stays local.
9.8KUpdated 5 months agoMIT
macOS · Windows · Linux#Batch processing#GGUF#Hugging Face integration
PowerInfer is a local LLM inference engine for developers and researchers who want to run large models on a PC with a consumer GPU. It splits work between the CPU and GPU to reduce GPU memory demands and data transfers. The code is open source under the MIT license.
62.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection
GPT-SoVITS is a local text-to-speech and voice cloning tool. It can generate speech from a short reference recording or fine-tune a model for a custom voice. The source code uses the MIT license.
3.4KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker#Code execution#MCP#Tool calling
Postgres MCP Pro is a self-hosted MCP server that gives AI assistants access to PostgreSQL alongside tools for diagnosing slow queries and testing index recommendations. It's for developers and database maintainers who want their assistant to work with the database's actual schema, query statistics and execution plans. It runs through Docker or Python and is open source under the MIT license.
541Updated 7 months agoGPL-3.0
macOS · Windows · Linux · iOS · Android#Multimodal input#Ollama integration
Reins is an open-source chat app for people using self-hosted LLMs through Ollama. It runs on iOS, Android, macOS, Linux and Windows, giving Ollama users a mobile and desktop interface for experimenting with models. Reins is the client; Ollama provides the model backend.
21.1KUpdated 4 days ago
macOS · Windows · Linux · Docker#ONNX#Voice conversion
Voice Changer (w-okada), also called VCClient, converts your voice as you speak using AI voice models. It's for people who want live voice conversion on their own computer, including those recording gaming commentary while running demanding software. Processing can stay local.
14.1KUpdated 3 years ago
macOS · Windows · Linux · Web#Multimodal input#Works offline
SadTalker generates talking head videos from a single portrait and an audio recording. The project states an Apache 2.0 license and removal of its earlier noncommercial restriction. It runs locally on Windows, Linux and macOS, and suits creators who want to animate a face without recording a person on camera. Its animation includes facial expressions and head movement, with examples covering speech and singing in different languages.
295Updated 2 days agoApache-2.0
Linux · Docker#OpenAI-compatible API#Wake word detection#Works offline
OpenVoiceOS is a free, open-source voice AI platform for people building their own smart speakers or adding voice control to devices. Its core builds on a fork of MycroftAI/mycroft-core, and most classic Mycroft skills also work with it. The project uses the Apache 2.0 license, which permits personal and commercial use.