9.6KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input
Xinference serves language, speech and multimodal models through a shared API on your own computer or servers. It's an open source platform under Apache 2.0 for developers and researchers who want to build applications around models they host. You can also deploy it on cloud infrastructure.
8.2KUpdated 2 weeks agoGPL-3.0
macOS#Image-to-image#Inpainting#Visual workflows
Dream Textures is an open-source Blender add-on for artists who want to generate images and apply AI textures to 3D work inside Blender. It runs Stable Diffusion on your own machine and connects image generation to scene depth, texture projection, and animation rendering.
5.8KUpdated 2 days agoMIT
macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend
llama-swap is a self-hosted proxy for people running several AI models on their own hardware. It starts the model server a request needs and swaps out another when necessary, so you don't have to keep every model loaded or manage separate API connections in your apps.
33.5KUpdated 2 days agoMIT
macOS · Windows · Linux · Web#GGUF#ONNX
Netron displays neural network and machine learning models as visual graphs. Developers and researchers can inspect a model's structure through desktop apps for macOS, Windows and Linux, or through the browser viewer at netron.app.
77.4KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#OpenAI-compatible API
GPT4All is a local AI chatbot for people who want to run language models on their own desktop or laptop and keep conversations on their machine. Its LocalDocs feature lets you ask questions about your own documents without sending them to a cloud service. It suits developers, teams and individuals who want control over their models and data.
40.1KUpdated 3 weeks agoApache-2.0
macOS · Linux · Web#Batch processing#llama.cpp backend#Multilingual
Marker is a local document converter for developers and teams turning PDFs, scans and Office files into structured text. It preserves tables, equations and page structure for document processing and AI workflows. Its pipeline reads embedded PDF text and uses Surya OCR where text is missing or damaged, rather than reading every page through a vision model.
8.2KUpdated 4 months agoApache-2.0
macOS · Windows · Linux · Web#Semantic search
sqlite-vec adds vector storage and similarity search to SQLite, so developers can keep embeddings alongside application data in a local database. It's for applications that need to find related items by vector distance without running a separate vector database server. The extension is small, written in C and has no dependencies.
2.5KUpdated 2 days agoMIT
macOS · Linux#Hugging Face integration#Multilingual
LightEval is a Python toolkit from Hugging Face for evaluating LLMs running on your own hardware or through remote services. It's for developers and researchers comparing models, investigating failures, or testing performance on tasks relevant to their work. It can evaluate a model already loaded in memory as well as one served through an endpoint.
3.8KUpdated 2 days agoMIT
macOS · Windows · Linux · Web#Batch processing#Voice conversion
Applio is a local AI voice conversion suite for musicians making AI covers, streamers changing their voice live, and creators working with speech. It converts recordings or microphone input into another voice using community models or models you train yourself. Its software uses the MIT license.
7.2KUpdated 1 day agoMIT
macOS#Batch processing#Distributed execution#Hugging Face integration
MLX LM is an open-source Python package for generating text and fine-tuning language models locally on Apple Silicon Macs. Built on MLX, it suits developers and researchers who want to work with models through Python or a terminal, including adapting models to their own tasks. The package uses the MIT license.
7.7KUpdated 4 weeks agoMIT
macOS · Windows · Linux#Batch processing#Multilingual#Ollama integration
Vibe is an open source desktop app for people who need transcripts or subtitles without uploading their recordings to a transcription service. It runs on macOS, Windows and Linux under the MIT license. Audio transcription works fully offline, with processing on your own computer.
1.5KUpdated 4 months agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API
Chatterbox TTS Server runs Resemble AI's speech models on your own computer or server, with a browser interface and an OpenAI-compatible API. It's for people producing narration and audiobooks, or developers adding speech to voice agents and other apps. The project is open source under the MIT license.
2.8KUpdated 2 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#Git integration#Human approval
Vexa is a meeting bot and transcription API for developers building meeting features and teams feeding calls into AI agents. Bots join Google Meet, Microsoft Teams and Zoom, then send live transcripts with speaker labels to your application or agent. You can host the full platform on your own infrastructure or use Vexa's cloud service.
21.7KUpdated 1 day agoApache-2.0
macOS#Distributed execution#Hugging Face integration#LoRA
PEFT is an open-source Python library for developers who want to adapt pretrained models on their own hardware with less compute and storage than full fine-tuning requires. It trains a small subset of parameters, often through adapters, while leaving the base model intact. It's licensed under Apache 2.0.
36.9KUpdated 2 years agoBSD-3-Clause
macOS · Windows · Linux#Batch processing
Real-ESRGAN is a local AI upscaler and restoration toolkit for people enlarging images or improving video, with dedicated models for anime illustrations and animation. It builds on ESRGAN and uses models trained entirely on synthetic data to address degraded images. The code is open source under the BSD 3-Clause license.
1KUpdated 1 month agoMIT
macOS
Practical-RIFE is a local AI frame interpolation tool for engineers and developers working with video. It generates intermediate frames to increase frame rates, with models intended for ordinary footage, animation, and post-processing videos made by diffusion models. The Python project builds on RIFE and SAFA, with an emphasis on how the output looks rather than improvements in numerical image-quality scores alone.
2.2KUpdated 3 days agoMIT
macOS · Windows · Linux#Batch processing#GGUF#Guardrails
node-llama-cpp is an open source library for developers adding local LLM inference to JavaScript and TypeScript applications. It connects Node.js, Bun and Electron to llama.cpp, running GGUF models on your own machine. Its MIT license allows use in commercial projects.
6KUpdated 2 months agoGPL-3.0
macOS · Windows · Linux#Batch processing#ONNX#Visual workflows
chaiNNer is a desktop image processing editor for people who want AI upscaling and repeatable editing workflows without writing scripts. It runs models locally on Windows, macOS, and Linux. Connected nodes let you combine model processing with ordinary image edits in the same workflow.
9.4KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · iOS · Android · Web#Batch processing#LLM tracing#Structured output
BAML is a programming language for developers building AI agents, with typed model calls and local tracing built into the language. It runs standalone on macOS, Linux and Windows, or alongside an existing application. The language is open source under Apache 2.0, and works offline.
19.9KUpdated 8 months agoMIT
macOS · Windows · Linux · Docker · Web#Code execution#Git integration#LM Studio integration
Bolt.diy is a self-hosted AI coding assistant for people building full-stack Node.js web apps who want to choose their own model provider. It's a fork of Bolt.new, with a browser workspace and an Electron desktop app for macOS, Windows, and Linux. You can also run it through Docker.
631Updated 2 years agoMIT
macOS · Windows · Linux · Web · Browser Extension#Batch processing#Multilingual#Quantization
TranslateLocally is an open source machine translation app for Windows, macOS and Linux. It's for people who want to translate text on their own device, including those handling material they don't want to send to a remote translation service. The desktop app has an MIT license and uses Marian and Bergamot translation models.
8.2KUpdated 4 weeks agoMIT
macOS · Windows · Linux#Code execution
Pinokio is a local launcher for people who want to try AI apps on their own computers. It pairs a catalog of community projects with scripts that handle the commands each app needs. Pinokio runs on Windows, macOS, and Linux, and its code is open source under the MIT license.
29.8KUpdated 1 day agoMIT
macOS · Windows · Linux#Code execution#Guardrails#Human approval
OpenAI Agents SDK is an open-source Python framework for developers building AI apps that need to use tools, delegate tasks, or work across multiple steps. Its runtime manages agent turns and conversation state while letting developers express workflows in ordinary Python. It uses the MIT license.
6KUpdated 3 months agoApache-2.0
macOS · iOS#Multimodal input#Ollama integration#Works offline
Enchanted is a ChatGPT alternative for macOS, iOS and visionOS that connects to models on your own Ollama server. It's for people who want an Apple app for chatting with privately hosted models such as Llama 2, Mistral, Vicuna and Starling. You supply the model server. The app is open source under the Apache License 2.0.
2.3KUpdated 4 months agoMPL-2.0
macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning
Coqui TTS (idiap fork) is a local text-to-speech library for developers and speech researchers who want pretrained voices or tools to train their own models. It builds on coqui-ai/TTS, continuing the original unmaintained project. The Python toolkit is open source under the Mozilla Public License 2.0 (MPL-2.0).
10.9KUpdated 1 day agoApache-2.0
macOS · Windows · Linux#Hugging Face integration#Multimodal input#ONNX
OpenVINO is an Apache 2.0 licensed toolkit for developers who want to run AI models locally or serve them on their own infrastructure. It converts and optimizes models for inference, with support for x86 and ARM CPUs, Intel integrated and discrete GPUs, and Intel NPUs. Its runtime works on Linux, Windows and macOS.
46.3KUpdated 1 day agoApache-2.0
macOS · Linux#Hybrid search#Semantic search
Milvus is an open-source vector database for developers building RAG applications, image search and recommendation systems. It stores embeddings alongside metadata so applications can retrieve related text, images or multimodal data. You can run it on your own hardware, from a laptop prototype to a distributed production cluster.
21.9KUpdated 7 months agoAGPL-3.0
macOS · Windows · Linux · Docker
Video2X runs AI video upscaling and frame interpolation on your own hardware. It's for people who want to increase a video's resolution or generate extra frames for smoother motion, including those working with anime footage. The project is open source under AGPL-3.0.
5.1KUpdated 22 hours ago
macOS · Windows · Linux#Git integration#MCP#Multi-agent workflows
Kiln is a desktop workbench for teams building AI applications on macOS, Windows and Linux. It keeps a task and its dataset together across evaluation, prompt optimization, RAG and fine-tuning, so teams can compare changes against the same examples. Engineers, data scientists, QA staff and subject matter experts can contribute through the app.
23.2KUpdated 1 day ago
macOS · Windows · Linux · Docker#Semantic search
pgvector adds vector storage and similarity search to Postgres, so developers can keep embeddings alongside application records in a self-hosted database. It suits applications that need to find similar items while retaining SQL queries, joins and transactional guarantees. It runs on Linux, macOS and Windows, with Docker also supported.
14.4KUpdated 1 day agoMIT
macOS · Windows · Linux#Batch processing#LM Studio integration#Multilingual
Subtitle Edit is an MIT-licensed subtitle editor for Windows, macOS and Linux. It's for people creating captions, translating dialogue or fixing subtitles that don't match the video. Its core editing, conversion and video playback work offline on your device, with optional AI tools for transcription and translation.
21.4KUpdated 3 weeks agoApache-2.0
macOS · Web#Batch processing#llama.cpp backend#Multilingual
Surya is a local OCR toolkit for developers extracting text and structure from PDFs and document images. It combines text recognition, layout analysis and table recognition in one vision-language model, so results retain page structure and reading order rather than just the words.
4.9KUpdated 7 months agoApache-2.0
macOS#LoRA#Multilingual
ACE-Step is an open-source music generation model for musicians, producers and developers who want to create and edit music on their own hardware. It generates songs with vocals or instrumental tracks from text descriptions and supplied lyrics. You can choose the duration and describe the sound with genre tags, longer prompts or a scene description.
77KUpdated 3 months agoAGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#MCP#Multi-user access#RAG
AppFlowy is a self-hosted Notion alternative for teams that want project management, shared documents, and AI in an environment they control. Its self-hosted enterprise offering can run on premises, in your own cloud, or in an air-gapped environment. Self-hosted LLMs and embedding models let you keep AI processing and workspace data inside your infrastructure, with no vendor access to your instance.
11KUpdated 3 days ago
macOS · Windows · Linux · Docker
nvtop is a terminal monitor that shows activity across multiple GPUs and accelerators in an interface familiar to htop users. It's useful for people running local LLM workloads or managing compute servers who want to see which processes are using their hardware and how much GPU memory they consume. It runs locally on Linux and is open source under GPLv3 or later.
21.1KUpdated 2 days agoApache-2.0
macOS · Web#GGUF#Hugging Face integration#Multilingual
Candle is a Rust machine learning framework for developers who want to embed local AI in applications or deploy models on their own servers. It produces lightweight binaries that don't need Python in production, making it a candidate for serverless inference where a large runtime can slow startup. Its API uses tensor operations familiar to PyTorch developers.