2.1KUpdated 7 hours agoAGPL-3.0
macOS · Windows · Linux#Agent Skills#MCP#Multilingual
OpenChatCut is a local-first AI video editor for creators who want conversational editing with control over the finished cut. AI changes become editable clips, captions, effects and audio tracks in the same project you can adjust manually. It's a free, open-source ChatCut alternative under AGPL-3.0, with a desktop app for macOS, Windows and Linux.
407Updated 2 days agoMIT
macOS#MLX#Multimodal input#OpenAI-compatible API
Slotstream runs Qwen3.8-Flash-Next on Apple Silicon Macs that don't have enough RAM to hold the whole model. It's aimed at people with 16 to 64 GB of memory who want local chat, image questions or a model backend for coding agents. Most model weights stay on the SSD, while frequently used expert networks stay in memory. The full model remains available.
2.4KUpdated 3 days agoAGPL-3.0
macOS · Windows · Web#Code execution#MCP#Multimodal input
PenEcho is an open-source workspace that puts handwritten notes, diagrams and AI-generated widgets on an infinite canvas. It runs on your computer as a macOS or Windows app, or through Node.js with a browser interface. It's for people who want to work through technical ideas visually and give AI feedback by drawing directly on the result.
3.1KUpdated 1 week agoMIT
macOS · Windows · Linux#Agent Skills#Multimodal input#Tool calling
Phone Harness connects an AI agent to a real iPhone or Android so it can use apps through their screens. It's for developers who want Claude Code, Codex, or another agent to carry out phone tasks without a dedicated app API. The software is open source under the MIT license, and it doesn't require a jailbreak or an app installed on the phone.
544Updated 14 hours agoMIT
macOS · iOS · Web#Code execution#Distributed execution#Hugging Face integration
Pooled runs a single open model across browser tabs on laptops, desktops and phones, combining their memory when the model won't fit on one device. It's for people who want local AI chat or a coding assistant using hardware they already have. It's open source under the MIT license and requires no account or per-device installation.
742Updated 4 hours agoAGPL-3.0
Web#MCP#Tool calling#Works offline
CozyClay is a local 3D previsualization studio for filmmakers and AI video creators who want to plan framing and movement before generating a finished clip. You build a rough scene, pose characters, and edit camera moves and cuts on a timeline. Those shots become visual references for models such as Seedance, Kling and Veo, so you can specify the camera through a clip rather than relying on text alone.
2.4KUpdated 1 day agoApache-2.0
macOS · Windows · Linux#Agent Skills#Git integration#MCP
ripwire helps AI coding agents find relevant code and assess changes without filling their context with whole files. It runs offline on your machine as a self-contained binary, with a CLI and an optional MCP server. It's for developers who want their agents to spend less context on repository research and check the effects of their edits.
1.1KUpdated 2 weeks agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Home Assistant integration#MCP#Streaming inference
Tel-Agent is a self-hosted AI receptionist for businesses that want phone calls and customer messages handled in one place. It runs on your own hardware and keeps recordings and searchable transcripts on your machine. It's free and open source under the AGPL-3.0 license.
1.1KUpdated 8 hours agoMIT
macOS · Windows · Linux#Agent Skills#MCP#Persistent memory
deja-vu gives coding assistants a shared memory of work already recorded on your machine. It searches sessions from before you installed it, so developers can recover an old fix or carry context between Claude Code, Codex CLI, Cursor and opencode without starting a separate collection of notes.
9.7KUpdated 1 day ago
macOS · Windows · Linux · Docker · Web#Batch processing#ControlNet#GGUF
Wan2GP brings video, image, music and speech generation to your own computer, with particular attention to GPUs with limited memory. It's for creators who want several media models in one browser interface. The project builds on Wan-Video/Wan2.1.
8.5KUpdated 23 hours agoApache-2.0
Web#LLM tracing#MCP#Multimodal input
Bifrost is a self-hosted AI gateway for developers and teams whose applications use multiple model providers. It puts Ollama, custom model deployments, and cloud services behind one OpenAI-compatible API, so applications can switch models without maintaining a separate integration for each provider.
16KUpdated 22 hours agoBSD-2-Clause
#LLM tracing#Multi-agent workflows#Multimodal input
Pipecat is a Python framework for developers building conversational AI agents that handle speech, video, text, and images. You can run it on your own machine or servers, wherever Python runs. It's open source under the BSD 2-Clause license.
29.8KUpdated 1 day agoApache-2.0
macOS · Docker · Web#Code execution#LLM tracing#Multi-user access
Sim is an AI agent workspace for teams that need to build automations and control how people use them across an organization. You can self-host it on your own server or cloud with Docker or Kubernetes. Its open-source code uses the Apache 2.0 license.
28.4KUpdated 21 hours ago
Web#Human approval#LLM tracing#MCP
Mastra is a TypeScript framework for developers building AI agents and applications on their own servers or inside existing web apps. Its server runs locally or as a standalone deployment; Mastra Cloud provides a hosted alternative. Model routing connects to providers such as OpenAI, Anthropic, and Gemini, so running the framework locally doesn't keep those model requests on your machine.
390.8KUpdated 19 hours ago
macOS · Windows · Linux · iOS · Android · Docker#Multi-user access#Persistent memory#Tool calling
OpenClaw is a self-hosted AI assistant that connects your own computer to the messaging services you already use. It's for people who want a personal assistant on their laptop, or teams that want to run a shared assistant on their own hardware. The project is open source under the MIT license.
155.4KUpdated 1 day agoMIT
macOS · Windows · Docker · Web#LLM tracing#MCP#Multi-agent workflows
Langflow is a visual builder for developers creating AI agents and retrieval-augmented generation (RAG) applications. You can run it locally or on your own server, with Docker support and desktop apps for Windows and macOS. The open-source software uses the MIT license. A hosted cloud offering provides a separate deployment option.
32.3KUpdated 1 day ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
93KUpdated 2 hours agoApache-2.0
macOS · Docker#Batch processing#Distributed execution#GGUF
vLLM is an open source engine for serving large language models on hardware you control. It suits developers and teams that need to handle many requests through an API while making efficient use of memory and compute. It's licensed under Apache 2.0 and can run with GPUs or on a CPU.
29.8KUpdated 4 days agoApache-2.0
Docker · Web#MCP#Multi-agent workflows#Ollama integration
GPT Researcher is a self-hosted AI research agent for people who need detailed, cited reports drawn from the web, their own documents, or both. You choose the language model and search provider, and can run the software on your own server or through Docker. It's open source under Apache 2.0.
157.6KUpdated 3 hours ago
Docker · Web#Code execution#MCP#OpenAI-compatible API
Dify is a source-available platform for teams building AI agents and apps on a visual canvas. Its Community Edition runs on your own server with Docker. Dify also offers a hosted cloud service, while Enterprise deployments can run in a VPC or on a self-hosted server. The Community Edition uses a custom Apache 2.0 derivative license.
28.4KUpdated 1 day ago
macOS · Windows · Linux · Android#llama.cpp backend#LM Studio integration#MCP
Crush is a terminal AI coding assistant for developers who want to work with their code and tools through a model they choose. The app runs on your machine, while model requests go to your selected local backend or cloud provider. It supports macOS, Linux, Windows (PowerShell and WSL), Android, FreeBSD, OpenBSD and NetBSD.
42.4KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#Human approval#LLM tracing
Agno is a Python framework and runtime for developers building customer-facing or internal AI agents. You can run its agent platform locally with Docker, on your own servers or in your cloud. The open-source framework uses the Apache 2.0 license, and the platform keeps sessions, memory, knowledge and traces in your database.
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
130KUpdated 1 hour agoMIT
Web#Code execution#GGUF#Hugging Face integration
llama.cpp runs language models on your own hardware and can serve them from a machine you control. It’s an MIT-licensed, open source inference engine for people building local AI apps, running a private model server, or using a model directly from the command line. It supports vision-language models too.
24.8KUpdated 25 minutes ago
Docker · Web#Code execution#Human approval#MCP
Activepieces brings AI agents and workflow automation into one workspace for teams that need to connect their apps and internal data. It runs on your own servers with Docker or Kubernetes, or as a managed cloud service. An isolated deployment can keep platform data on its own infrastructure; connected apps and model providers may still require external access. Its core and app integration pieces are MIT-licensed.
29.9KUpdated 1 day ago
Docker · JetBrains#Code execution#MCP#Persistent memory
Serena is a locally run MCP toolkit that gives AI coding agents access to code structure: functions, classes and the references between them. It's for developers who want their agent to find and change specific parts of a codebase without reading whole files or relying on text matching. It requires an LLM client and works with Claude Code, Codex, OpenCode, Cursor and OpenWebUI.
27.3KUpdated 21 hours agoMIT
macOS · Windows · Linux#Code execution#MCP#Tool calling
Cua gives AI agents access to computers they can inspect and operate, with tools for desktop automation, local virtual machines, and hosted fleets. It's for developers building agents that work across native apps and browsers, or evaluating how well those agents complete computer tasks. You bring the agent and model.
62.5KUpdated 2 days agoMIT
#MCP#RAG#Tool calling
Context7 brings current library documentation and code examples into an AI coding assistant's context. It's for developers who want answers grounded in the libraries and versions they're actually using, rather than an assistant's older training data. It works with Claude Code, Codex, Cursor, Devin Desktop and Antigravity.
20.3KUpdated 2 hours agoMIT
#Human approval#LLM tracing#MCP
Pydantic AI is a Python SDK for developers building AI agents into their own applications. Its main draw is Pydantic validation across agent tools and results, so an agent can return structured data that application code can check and use. The SDK is MIT licensed.
127.4KUpdated 12 minutes agoApache-2.0
macOS · Windows · Linux#Code execution#Git integration#Human approval
Codex CLI is an open source coding assistant that runs on your computer and works with local repositories. It's for developers who want to explore code, make changes, and use their existing development tools from the terminal. It runs on macOS, Windows, and Linux under the Apache 2.0 license.
250.1KUpdated 19 hours agoMIT
macOS · Windows · Linux · Android · Docker#Agent Skills#Code execution#Multi-agent workflows
Hermes Agent is a self-hosted AI agent from Nous Research for people who want an assistant that remembers past work and develops reusable skills. It creates skills after complex tasks, revises them through use, and searches earlier conversations to recover context across sessions. The project is open source under the MIT license.
75.2KUpdated 2 days agoApache-2.0
Web#LoRA#Multimodal input#OpenAI-compatible API
LLaMA-Factory is an open-source framework for developers and researchers who want to adapt language and multimodal models on their own hardware. It brings training and inference into one toolkit, with support for LLaMA, Qwen3, Qwen3-VL, DeepSeek, Gemma, Mistral and LLaVA. Its license is Apache 2.0.
52.3KUpdated 1 hour agoAGPL-3.0
macOS · Windows · Linux#LM Studio integration#MCP#Multilingual
Cherry Studio is a free, open source desktop app for people who use several AI models and want their conversations in one place. It runs on Windows, macOS and Linux. Chats, settings and knowledge base files are stored on your device.
59.9KUpdated 1 hour ago
#Guardrails#MCP#Multi-user access
LiteLLM gives platform teams one place to manage access to LLMs across providers. Its self-hosted AI gateway puts cloud services and internal or locally hosted models behind an OpenAI-compatible API, so applications can change models without changing their integration. Developers can also use its Python SDK directly.
21.8KUpdated 21 hours ago
macOS · Windows · Linux#MCP#Multimodal input#Ollama integration
Screenpipe records screen activity and audio as searchable local history that AI agents can use as context. It's for people who want to recall past work and teams that want agents to draft follow-ups or update work records using what actually happened. Raw history stays on your device by default.
182KUpdated 17 hours agoMIT
macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#Multimodal input
Ollama runs language models on your own computer or server. It provides a command-line runner and a local API for people building AI applications or connecting existing tools to models they host themselves. The software is distributed under the MIT license.