3.9KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#Multimodal input
Riffusion is a Python library for generating music and audio on your own hardware using Stable Diffusion. It's for developers and musicians who want to experiment with text-driven sound generation or build it into an app. The hobby project is no longer actively maintained.
10.3KUpdated 1 year agoMIT
macOS · Windows · Linux#Multimodal input#Ollama integration
Self-Operating Computer lets a vision-capable AI model control a desktop by reading the screen and choosing mouse and keyboard actions to carry out a goal. It's a Python framework for developers and researchers exploring AI agents that work through application interfaces. The code uses the MIT license.
520Updated 13 hours ago
macOS · Windows · Linux#Visual workflows#Works offline
Comfy Desktop installs and launches ComfyUI on your computer, taking care of the Python environment and dependencies. You can manage several independent ComfyUI instances, each with its own version, custom nodes and settings. It suits people who want to run local image and media workflows without setting up every environment by hand.
4.7KUpdated 1 month agoMIT
macOS · Windows · Linux#Visual workflows
Rivet is a desktop visual programming environment for developers building AI agents and applications with complex LLM workflows. Its editor runs on macOS, Windows and Linux, and its TypeScript library executes the resulting graphs inside your own application. The project uses the MIT license.
7.3KUpdated 17 hours agoApache-2.0
Linux · Docker · Web#Multi-user access
Civitai is a self-hostable platform for sharing AI image models and artwork. Its Apache-2.0 code provides accounts, model uploads, browsing and comments. The public Civitai website also operates a hosted image generator, which is separate from the runnable local platform.
2.9KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux · Web#Semantic search#Works offline
OpenRecall records your screen at regular intervals and makes that history searchable with local AI. It's a free, open-source alternative to Microsoft's Windows Recall and Rewind.ai for people who want to find something they previously saw on their computer. It runs on Windows, macOS and Linux, with a browser interface served from your own machine.
19.6KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker · Web#Ollama integration#Web search
Devika is a self-hosted AI coding agent for developers who want to give a software task in plain language and have an agent plan the work, research it and write code. Modeled after Cognition AI's Devin, it runs on your own machine with a browser interface and supports local LLMs through Ollama. It's open source under the MIT license.
278Updated 1 day agoAGPL-3.0
macOS · Windows · Linux · Android · Web#Works offline
Writeopia is a writing and note-taking app for people who want AI assistance while keeping their documents on their own computer. It runs on Windows, Linux and macOS, with offline access to locally stored notes and a choice of AI models that run on your device. It's aimed at writers, developers and researchers working on drafts, documentation or personal notes.
1.8KUpdated 2 years ago
macOS · Windows · Linux#Batch processing#Hugging Face integration
Stable Fast 3D turns a single object image into a textured 3D mesh on your own hardware. It's aimed at game and VR developers, designers and people creating product models for e-commerce. Built on TripoSR, it uses a retrained model designed to produce meshes and textures suitable for use in games and other 3D projects.
2.6KUpdated 2 years ago
macOS · Linux · Web#Batch processing#Hugging Face integration
AudioLDM 2 generates sound effects, music and speech on your own hardware. It's a Python tool for people experimenting with synthetic audio, including sound designers and researchers who want to work with pretrained models. A Gradio browser interface and command-line tools provide access to local generation; a hosted Hugging Face demo is also available.
4.2KUpdated 1 year agoApache-2.0
Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration
LMQL is a programming language for developers who need model calls and ordinary Python logic in the same program. It lets you define rules for generated text, including types, length limits, allowed answers and stopping phrases. Those rules apply during generation, so you can constrain intermediate responses as well as the final output.
15.9KUpdated 2 months agoApache-2.0
macOS · Windows · Linux · VS Code#Git integration
DVC connects data and model versions to the code in your Git repository, so you can reproduce a machine learning experiment with the inputs it used. It's a free, open-source tool under Apache 2.0 for individual data scientists and small projects. It runs on macOS, Windows and Linux.
2.3KUpdated 10 months agoApache-2.0
macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input
DiffRhythm is a local AI music generation model for musicians, developers and researchers who want to create full-length songs on their own hardware. It uses latent diffusion to generate songs with vocals and accompaniment, and can also produce instrumental music. The full model supports songs up to 4 minutes and 45 seconds.
38.1KUpdated 1 day agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Code execution#MCP#Single sign-on
Trilium Notes is a local-first note-taking app for people building a large personal knowledge base. It runs on Windows, macOS and Linux, or on your own server through Docker, with browser access and a mobile web interface. It's free and open source under AGPL-3.0.
47.9KUpdated 1 day agoGPL-2.0
Linux · Web#LLM tracing#Multi-user access#Structured output
Discourse AI is the official AI plugin bundled with Discourse. Forum administrators can enable its features independently, including an AI bot, semantic search, topic and chat summaries, spam detection and writing assistance. It runs inside a Discourse community, which you can host on your own server.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
2.7KUpdated 1 week agoApache-2.0
Linux · Docker#Hugging Face integration#Quantization
Intel Neural Compressor is a Python library for developers compressing AI models for deployment on their own hardware or servers. It supports local LLM work as well as other deep learning models, with particular attention to Intel CPUs, GPUs and Gaudi accelerators. It's open source under the Apache 2.0 license.
2.3KUpdated 1 year agoMIT
Linux#Batch processing#GGUF#Hugging Face integration
AutoAWQ is a Python library for developers who want to compress and run LLMs on their own hardware using 4-bit Activation-aware Weight Quantization (AWQ). The project is archived and no longer maintained. It's open source under the MIT license. It installs as a Python package, with optional kernel or Intel CPU dependencies.
5.2KUpdated 4 days agoApache-2.0
Linux · Docker · Web#Hugging Face integration#LoRA#Quantization
H2O LLM Studio is a self-hosted tool for teams that want to adapt language models to their own datasets without writing training code. Its browser interface brings training experiments, evaluation, and model testing into one place. The project is open source under Apache 2.0.
14KUpdated 5 months ago
macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP
Open-LLM-VTuber is a local AI companion for people who want a character they can talk to, with a Live2D avatar that responds through speech and expressions. It runs on Windows, macOS and Linux through web and desktop clients. With local models for speech and language processing, it works fully offline and keeps conversations on your device. Cloud APIs are optional alternatives that send the corresponding processing to external services.
7.2KUpdated 2 years agoApache-2.0
macOS · Linux · Docker · Web#Multilingual#Voice cloning
Zonos is an open-source text-to-speech model for people who want to generate speech and clone voices on their own hardware. It can match a speaker from a short reference recording, with controls for delivery and emotion. The code uses the Apache 2.0 license.
9.4KUpdated 1 day agoMIT
macOS · Windows · Linux · iOS · Android#LM Studio integration#MCP#Ollama integration
Anarlog, formerly Hyprnote, is a desktop AI meeting notetaker for people who want to keep private conversations on their own hardware. It captures audio from your device without adding a bot to the call and stays hidden during screen sharing. The app runs on macOS, Windows and Linux; its community application is open source under the MIT license.
953Updated 3 weeks agoMIT
macOS · Windows · Linux#Ollama integration
Ollama Grid Search is a desktop app for comparing LLM responses across models, prompts and inference settings. It runs on macOS, Windows and Linux, and suits developers or anyone choosing a model and prompt combination for a particular task. You can inspect the responses together rather than repeat each test by hand.
2.3KUpdated 4 months agoMPL-2.0
macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning
XTTS v2 generates speech from text using a reference voice recording or a preset speaker. It runs locally through Coqui TTS and suits developers building speech into apps, as well as researchers who want to fine-tune a speech model on their own hardware.
12.3KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval
ShellGPT is an AI terminal assistant for developers and people who work with shell commands. It runs on Linux, macOS and Windows, turning plain-language requests into commands suited to your operating system and shell. You can review, explain or execute its suggestions, and its Bash and Zsh integrations put generated commands into the terminal input line for editing.
784Updated 4 months agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multi-user access#Multimodal input#Persistent memory
Agnai is a self-hosted AI roleplay chat app for people who want to create fictional characters and talk with them alone or in a group. A conversation can include multiple people and multiple bots. It builds on early work from Galatea-UI by PygmalionAI and uses the AGPL-3.0 open-source license.
1.7KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#Multilingual#Persistent memory
RisuAI is an open-source AI roleplay client for people who want to create characters, build fictional worlds and chat with several characters together. It runs on Windows, macOS, Linux, Android and in a browser. You can also host the web app yourself with Docker.
1.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux#Ollama integration#RAG#Works offline
tlm is an open source terminal assistant for people who want help writing shell commands and understanding unfamiliar ones. It runs models on your workstation through Ollama, so command assistance doesn't depend on a cloud service. It works offline and requires no API key or subscription.
4.5KUpdated 7 months agoMIT
macOS · Windows · Linux#MCP#Ollama integration#OpenAI-compatible API
mods is a command-line AI tool for people who want to ask questions about command output or use model responses in shell pipelines. It connects to local LLMs through LocalAI as well as cloud services. The project is archived and no longer maintained.
1.1KUpdated 3 weeks agoMPL-2.0
Linux#Ollama integration#RAG#Semantic search
chromem-go is a vector database that runs inside your Go application, so developers can add semantic search or retrieval augmented generation (RAG) without maintaining a separate database server. It stores text alongside embeddings and retrieves related documents for use in LLM answers. Its focus is ordinary application workloads rather than collections containing millions of documents.
5.4KUpdated 3 months ago
macOS · Windows · Linux#MCP#Multilingual#Ollama integration
5ire is a free desktop AI assistant that combines chat, a local document knowledge base and MCP tools. It runs on macOS, Windows and Linux, with Mac downloads for Apple Silicon and Intel. It's for people who want to use their own documents and external tools alongside conversations with local or cloud models.
4.8KUpdated 3 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
Lollms WebUI is a local, single-user AI interface for people who want text chat and media generation in one place. It runs on Windows, macOS and Linux, with Docker support, and lets writers, developers and other users choose models and task-specific personalities. It's free and open source under Apache 2.0. The project receives minimal maintenance.
33.3KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Ollama integration
Chatbot UI is a browser-based AI chat app for people who want to host their own interface and choose between local models and cloud providers. It connects to Ollama for models running on your hardware, alongside services such as OpenAI, Azure OpenAI, Anthropic and Google Gemini. The app is open source under the MIT license.
1.9KUpdated 3 weeks agoAGPL-3.0
macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration
Sonar is a self-hosted inference engine for developers and teams serving Hugging Face-compatible language and multimodal models on their own hardware. Based on vLLM, it adds model and quantization formats, sampling methods, and deployment features. It's open source under AGPL-3.0.
qualcomm/GenieXInference Libraries and Bindings
macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend
Nexa SDK is an on-device AI inference framework for developers building applications that process text, images or audio on users' hardware. It runs models locally across CPUs, GPUs and NPUs, with a shared interface for different backends. Its scope includes language and vision models, speech recognition, speech synthesis and image generation.
huggingface.coOCR and Document Scanning
Linux#Batch processing#Hugging Face integration#Multimodal input
Qwen2.5-VL is a vision-language model you can run on your own hardware to answer questions about images and video. It's aimed at developers building document processing tools, visual assistants and agents that interact with computer or phone screens. The instruction-tuned 7B model has Apache 2.0 licensing and works with Hugging Face Transformers, with weights available in Safetensors format.