14.2KUpdated 1 week agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
QAnything is a self-hosted knowledge base for people and teams who want to ask questions about their own documents, including collections that mix Chinese and English. It can answer in either language regardless of the document's language, and runs locally through Docker on Windows, macOS and Linux.
3.3KUpdated 2 months agoMIT
Windows · Linux · Docker · Web#Hugging Face integration#LoRA
FluxGym is a local web interface for training FLUX LoRAs on your own images, with support for GPUs with 12GB, 16GB or 20GB of VRAM. It's for people who want to customize an image model through a browser while retaining access to detailed training controls. It runs on Windows and Linux, with Docker support, and is open source under the MIT license.
253Updated 4 weeks agoMIT
macOS · Windows#MCP#Multilingual#Multimodal input
Raycast Ollama brings Ollama models into Raycast on macOS and Windows, for people who want AI chat and text assistance within their desktop launcher. It connects to Ollama on your own machine or a remote server you choose. The extension is open source under the MIT license, and local inference doesn't require an Ollama API key.
1.5KUpdated 2 years agoMIT
macOS · Windows · Browser Extension#Multimodal input#Ollama integration#RAG
Lumos is a Chrome extension for asking questions about web pages using models running on your own machine. It's for readers working through long discussions, product reviews, news articles or technical documentation who want summaries and answers tied to the material they're reading. Ollama handles inference locally, without a remote AI server.
857Updated 2 years agoAGPL-3.0
macOS · Windows · Linux · Docker#Multilingual#ONNX#OpenAI-compatible API
OpenedAI Speech is a self-hosted text-to-speech server for developers who want local speech generation in apps built around OpenAI's speech API. The project is archived and no longer maintained. It's open source under AGPL-3.0, and it generates audio on your own hardware without an OpenAI API key.
3.9KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#Multimodal input
Riffusion is a Python library for generating music and audio on your own hardware using Stable Diffusion. It's for developers and musicians who want to experiment with text-driven sound generation or build it into an app. The hobby project is no longer actively maintained.
10.3KUpdated 1 year agoMIT
macOS · Windows · Linux#Multimodal input#Ollama integration
Self-Operating Computer lets a vision-capable AI model control a desktop by reading the screen and choosing mouse and keyboard actions to carry out a goal. It's a Python framework for developers and researchers exploring AI agents that work through application interfaces. The code uses the MIT license.
7.7KUpdated 4 months agoBSD-3-Clause
Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
Verba is a self-hosted document chatbot for people who want to ask questions across their files and knowledge bases. The project is archived and no longer maintained. It uses Weaviate to find relevant passages and gives those passages to a language model to generate answers.
966Updated 9 months agoGPL-3.0
Windows#Batch processing#Image-to-image#Inpainting
NMKD Stable Diffusion GUI is a local AI image generator for people who want to create and edit images on a Windows PC. It combines Stable Diffusion generation with inpainting, LoRA training and image post-processing in a desktop interface. It's open source under GPL-3.0.
12.7KUpdated 2 months agoMIT
Windows · Web#Human approval#MCP#Multi-agent workflows
Open Deep Research is a self-hosted AI agent that searches for information and writes research reports. It's for developers and teams who want to choose their own models and research tools. The project is archived and no longer maintained.
13.1KUpdated 2 years agoApache-2.0
macOS · Windows#Batch processing#Hugging Face integration#Multilingual
insanely-fast-whisper is a command-line tool for people who want to transcribe audio on their own hardware, with a focus on processing long recordings quickly. It runs OpenAI's Whisper locally on NVIDIA GPUs or Apple Silicon Macs, including support for Windows with CUDA. The project is open source under the Apache 2.0 license.
520Updated 13 hours ago
macOS · Windows · Linux#Visual workflows#Works offline
Comfy Desktop installs and launches ComfyUI on your computer, taking care of the Python environment and dependencies. You can manage several independent ComfyUI instances, each with its own version, custom nodes and settings. It suits people who want to run local image and media workflows without setting up every environment by hand.
4.7KUpdated 1 month agoMIT
macOS · Windows · Linux#Visual workflows
Rivet is a desktop visual programming environment for developers building AI agents and applications with complex LLM workflows. Its editor runs on macOS, Windows and Linux, and its TypeScript library executes the resulting graphs inside your own application. The project uses the MIT license.
33.7KUpdated 4 months ago
Windows · Docker#Code execution#Human approval#Multi-agent workflows
GPT Pilot is a self-hosted AI coding assistant for developers who want an agent to build applications under their supervision. The project is no longer maintained. It uses the FSL-1.1-MIT source-available license, with restrictions on competing uses. It runs locally as a Python CLI and is the core technology behind the Pythagora VS Code extension.
55.1KUpdated 2 years agoMIT
Windows · Docker#Code execution#Multimodal input
gpt-engineer is a locally run coding assistant for developers who want to experiment with AI code generation and build their own agents. It can create software from plain-language descriptions or make requested changes to an existing codebase. The project is archived and no longer maintained. Its Python code is open source under the MIT license.
2.9KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux · Web#Semantic search#Works offline
OpenRecall records your screen at regular intervals and makes that history searchable with local AI. It's a free, open-source alternative to Microsoft's Windows Recall and Rewind.ai for people who want to find something they previously saw on their computer. It runs on Windows, macOS and Linux, with a browser interface served from your own machine.
19.6KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker · Web#Ollama integration#Web search
Devika is a self-hosted AI coding agent for developers who want to give a software task in plain language and have an agent plan the work, research it and write code. Modeled after Cognition AI's Devin, it runs on your own machine with a browser interface and supports local LLMs through Ollama. It's open source under the MIT license.
278Updated 1 day agoAGPL-3.0
macOS · Windows · Linux · Android · Web#Works offline
Writeopia is a writing and note-taking app for people who want AI assistance while keeping their documents on their own computer. It runs on Windows, Linux and macOS, with offline access to locally stored notes and a choice of AI models that run on your device. It's aimed at writers, developers and researchers working on drafts, documentation or personal notes.
hillnote.comAI Notes and Knowledge Bases
macOS · Windows · iOS · Android#MCP#Ollama integration#Works offline
Hillnote is a writing and planning workspace for people who want their notes on their own disk, with AI available inside the editor. It runs on Mac, Windows, iOS and Android. The editor, workspace and local AI work offline, and getting started doesn't require an account.
1.8KUpdated 2 years ago
macOS · Windows · Linux#Batch processing#Hugging Face integration
Stable Fast 3D turns a single object image into a textured 3D mesh on your own hardware. It's aimed at game and VR developers, designers and people creating product models for e-commerce. Built on TripoSR, it uses a retrained model designed to produce meshes and textures suitable for use in games and other 3D projects.
4.2KUpdated 1 year agoApache-2.0
Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration
LMQL is a programming language for developers who need model calls and ordinary Python logic in the same program. It lets you define rules for generated text, including types, length limits, allowed answers and stopping phrases. Those rules apply during generation, so you can constrain intermediate responses as well as the final output.
nurgo-software.comAI Workflow Automation
Windows#Code execution#Multi-agent workflows#Multimodal input
BrainSoup is a proprietary native Windows app for people who want custom AI agents to handle work on their desktop. You can give agents distinct roles and access to different data, then have them collaborate in shared chat rooms. Natural-language conversations guide their tasks and automations.
2.1KUpdated 2 years agoMIT
macOS · Windows · VS Code#Multilingual#Ollama integration#Quantization
Llama Coder is an open source VS Code extension for developers who want a self-hosted alternative to GitHub Copilot's code completion. It uses Ollama to run models on your own hardware, either on the computer you're coding on or on a separate machine. The extension has no telemetry or tracking.
15.9KUpdated 2 months agoApache-2.0
macOS · Windows · Linux · VS Code#Git integration
DVC connects data and model versions to the code in your Git repository, so you can reproduce a machine learning experiment with the inputs it used. It's a free, open-source tool under Apache 2.0 for individual data scientists and small projects. It runs on macOS, Windows and Linux.
2.3KUpdated 10 months agoApache-2.0
macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input
DiffRhythm is a local AI music generation model for musicians, developers and researchers who want to create full-length songs on their own hardware. It uses latent diffusion to generate songs with vocals and accompaniment, and can also produce instrumental music. The full model supports songs up to 4 minutes and 45 seconds.
63.5KUpdated 11 months agoMIT
macOS · Windows#Hugging Face integration
nanoGPT is a Python toolkit for developers and researchers who want to train GPT models on their own hardware or fine-tune existing GPT-2 checkpoints. Its author has deprecated the project and points readers to nanochat. The MIT-licensed code remains available for study and modification.
38.1KUpdated 1 day agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Code execution#MCP#Single sign-on
Trilium Notes is a local-first note-taking app for people building a large personal knowledge base. It runs on Windows, macOS and Linux, or on your own server through Docker, with browser access and a mobile web interface. It's free and open source under AGPL-3.0.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
5.1KUpdated 22 hours agoApache-2.0
Windows · Docker#Agent Skills#Hugging Face integration#ONNX
TensorRT Model Optimizer, called NVIDIA Model Optimizer or ModelOpt, is a Python library for developers preparing models for local or self-hosted inference. It reduces model size and memory use and can speed up inference through compression and other optimization techniques. It's open source under Apache 2.0.
14KUpdated 5 months ago
macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP
Open-LLM-VTuber is a local AI companion for people who want a character they can talk to, with a Live2D avatar that responds through speech and expressions. It runs on Windows, macOS and Linux through web and desktop clients. With local models for speech and language processing, it works fully offline and keeps conversations on your device. Cloud APIs are optional alternatives that send the corresponding processing to external services.
6.4KUpdated 3 years agoMIT
Windows#Hugging Face integration#Multilingual#Voice cloning
StyleTTS 2 is an open-source text-to-speech model for developers and speech researchers who want to generate expressive speech on their own hardware. It can choose a speaking style from the text without a reference recording, while its multispeaker model uses reference audio to reproduce a speaker's voice and delivery. The Python code uses PyTorch and carries the MIT license.
14.7KUpdated 1 year agoApache-2.0
Windows#Hugging Face integration#Multimodal input
Sesame CSM is an open-source speech generation model for developers and researchers building voice applications on their own hardware. It uses text and audio inputs to generate speech, with support for conversational context and different speakers. It's a model component for applications that need spoken output.
1.6KUpdated 1 year agoMIT
Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input
Amica is a locally runnable interface for talking with customizable 3D AI characters. It's for people who want an animated, voiced character as the face of their AI assistant, with a choice of local LLM backends or cloud services. The project builds on Pixiv's ChatVRM.
9.4KUpdated 1 day agoMIT
macOS · Windows · Linux · iOS · Android#LM Studio integration#MCP#Ollama integration
Anarlog, formerly Hyprnote, is a desktop AI meeting notetaker for people who want to keep private conversations on their own hardware. It captures audio from your device without adding a bot to the call and stays hidden during screen sharing. The app runs on macOS, Windows and Linux; its community application is open source under the MIT license.
953Updated 3 weeks agoMIT
macOS · Windows · Linux#Ollama integration
Ollama Grid Search is a desktop app for comparing LLM responses across models, prompts and inference settings. It runs on macOS, Windows and Linux, and suits developers or anyone choosing a model and prompt combination for a particular task. You can inspect the responses together rather than repeat each test by hand.
2.3KUpdated 4 months agoMPL-2.0
macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning
XTTS v2 generates speech from text using a reference voice recording or a preset speaker. It runs locally through Coqui TTS and suits developers building speech into apps, as well as researchers who want to fine-tune a speech model on their own hardware.