13.2KUpdated 1 day agoApache-2.0
#MCP#Ollama integration#RAG
LangChain4j is an Apache 2.0 open-source Java library for developers building chatbots, assistants and AI agents in JVM applications. It connects application code to local LLM backends such as Ollama as well as cloud providers such as OpenAI and Google Vertex AI. Where model requests go depends on the backend you choose.
4.4KUpdated 1 day agoMIT
#Human approval#Multi-agent workflows#Multimodal input
RubyLLM is an MIT-licensed AI framework for developers building Ruby and Rails applications with local or hosted models. Its shared API lets an application switch between Ollama, cloud providers such as Anthropic and OpenAI, and OpenAI-compatible endpoints without rewriting its model integration. The framework runs in your application; model processing happens at the local or hosted backend you choose.
28.6KUpdated 2 days agoMIT
Docker · Web · Browser Extension#Agent Skills#Code execution#MCP
Repomix turns a code repository into a single file that an AI assistant can read. It’s for developers who need to give Claude, ChatGPT, Gemini, or another model enough project context for code review or analysis. You can use its command line on your own machine, pack a repository through its website, or run it as an MCP server, including through Docker. It is open source under the MIT license.
3.9KUpdated 1 week agoApache-2.0
#Hugging Face integration#Multilingual#Tool calling
SmolLM3 is a 3B parameter language model from Hugging Face for developers and researchers who want to run an LLM on their own hardware. It comes as a base model and an instruction-tuned model for chat, reasoning and tool calling. Both run locally.
2.8KUpdated 2 weeks agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#Git integration#Human approval
Vexa is a meeting bot and transcription API for developers building meeting features and teams feeding calls into AI agents. Bots join Google Meet, Microsoft Teams and Zoom, then send live transcripts with speaker labels to your application or agent. You can host the full platform on your own infrastructure or use Vexa's cloud service.
1.4KUpdated 2 months agoMIT
#Multimodal input#Ollama integration#Streaming inference
OllamaSharp is a C# library for developers building .NET applications around Ollama. It connects to Ollama on your own machine or a remote server and covers the full Ollama API, including model management alongside chat and embeddings. It's open source under the MIT license.
18.2KUpdated 4 days agoApache-2.0
Windows#Agent Client Protocol#llama.cpp backend#Ollama integration
avante.nvim brings Cursor-style AI assistance into Neovim for developers who want to keep their editor and choose their own model backend. It answers questions about code, suggests changes and applies edits directly to source files. The plugin is open source under Apache 2.0.
4.8KUpdated 1 day ago
Docker · Web#Git integration#Human approval#LLM tracing
Agenta is an MIT-licensed, open source workspace for teams that want AI agents to handle tasks in chat and continue recurring work in the background. You can self-host it with Docker Compose or Helm to keep agents and workspace data on your infrastructure, or use Agenta Cloud as a hosted service.
22.9KUpdated 3 days agoGPL-3.0
Docker · Web#MCP#Multimodal input#RAG
MaxKB is a self-hosted AI agent platform for organizations building customer support bots, internal knowledge assistants and business automation. It combines answers grounded in company documents with workflows that can call functions and MCP tools. You can run it on your own server through Docker and use it through a browser.
2.2KUpdated 3 days agoMIT
macOS · Windows · Linux#Batch processing#GGUF#Guardrails
node-llama-cpp is an open source library for developers adding local LLM inference to JavaScript and TypeScript applications. It connects Node.js, Bun and Electron to llama.cpp, running GGUF models on your own machine. Its MIT license allows use in commercial projects.
84.5KUpdated 1 day agoBSD-3-Clause
Docker#Guardrails#MCP#Tool calling
Scrapling is a Python web scraping framework for developers collecting website data or giving AI agents access to web pages. Its adaptive parser can find previously selected elements after a site's layout changes, reducing the need to repair extraction rules. It's open source under the BSD-3-Clause license and runs on your own machine or in Docker.
1.3KUpdated 1 month ago
Web#Human approval#MCP#Multi-user access
Hexabot is a self-hosted AI workflow automation platform for teams building customer service assistants and business automations. It connects conversations to actions across websites, messaging platforms, social channels and custom entry points. You can run it on-premise or in your own cloud, with workflows, memory and customer data in infrastructure you control.
6.9KUpdated 22 hours agoApache-2.0
#Agent Client Protocol#Human approval#MCP
codecompanion.nvim brings model chat, inline code edits, and coding agents into Neovim. It's for developers who want AI assistance inside their editor, with a choice between local LLMs through Ollama and cloud providers such as Anthropic, OpenAI, and Google Gemini. The plugin runs in Neovim; where model processing happens depends on the backend you connect.
29.8KUpdated 1 day agoMIT
macOS · Windows · Linux#Code execution#Guardrails#Human approval
OpenAI Agents SDK is an open-source Python framework for developers building AI apps that need to use tools, delegate tasks, or work across multiple steps. Its runtime manages agent turns and conversation state while letting developers express workflows in ordinary Python. It uses the MIT license.
3.5KUpdated 6 days agoGPL-3.0
#Git integration#Human approval#llama.cpp backend
gptel brings LLM conversations into the Emacs buffers where you already write and work. It's for Emacs users who want AI help alongside their text or code, with a choice of local model servers and cloud providers. The client is open source under GPL-3.0.
8.2KUpdated 21 hours ago
#Distributed execution#Multimodal input#OpenAI-compatible API
NVIDIA Dynamo is a self-hosted inference framework for teams serving models across multiple GPUs or server nodes. It coordinates SGLang, TensorRT-LLM and vLLM, adding cluster-level scheduling and request routing above those engines. Its focus is large deployments where GPU capacity, response latency and repeated computation affect serving costs.
9.3KUpdated 2 months agoApache-2.0
#Multilingual#Multimodal input#RAG
PaperQA2 is an open source Python research assistant for people who need answers grounded in a collection of scientific papers. It searches documents on your machine and writes answers with in-text citations, including page references. Researchers can use it to summarize findings or check for contradictions across papers, while developers can build it into their own research tools.
5.1KUpdated 22 hours ago
macOS · Windows · Linux#Git integration#MCP#Multi-agent workflows
Kiln is a desktop workbench for teams building AI applications on macOS, Windows and Linux. It keeps a task and its dataset together across evaluation, prompt optimization, RAG and fine-tuning, so teams can compare changes against the same examples. Engineers, data scientists, QA staff and subject matter experts can contribute through the app.
4.4KUpdated 2 days agoBSD-3-Clause
Web#Agent Client Protocol#Code execution#Human approval
Jupyter AI connects AI agents to notebooks through a chat interface inside JupyterLab. It's for people who work in computational notebooks and want an agent to help write code, debug cells, or run notebook work without moving the conversation to a separate app. The extension runs within JupyterLab; the choice of agent determines the AI service it connects to.
19.2KUpdated 2 weeks agoApache-2.0
Web · Browser Extension#OpenAI-compatible API#Streaming inference#Structured output
WebLLM runs language models directly in a user's browser, using WebGPU for GPU acceleration. It's an open-source engine for developers building web-based AI assistants and Chrome extensions that process prompts on the user's device rather than an inference server. The project uses the Apache 2.0 license.
82.9KUpdated 4 hours ago
Linux · Docker · Web#MCP#Multi-agent workflows#Multi-user access
LobeHub is an AI agent workspace for people who want to assign work to several assistants and keep that work organized. It offers a hosted service and a Docker-based self-hosted version for a private device or server. A Linux download is also available.
26.6KUpdated 1 day agoApache-2.0
Docker#Guardrails#Hugging Face integration#Hybrid search
Haystack is a Python framework for developers building self-hosted AI agents, document search, and apps that answer questions using their own data. Its modular pipelines let teams control which information reaches a model and inspect how retrieval, memory, tools, and generation contribute to an answer. It's open source under Apache 2.0.
7.2KUpdated 24 hours ago
Docker#Guardrails#LLM tracing#Tool calling
NeMo Guardrails is an open-source Python toolkit for developers who need control over how an AI assistant responds and uses tools. It runs within your application or as a self-hosted server, including in Docker. The library uses the Apache 2.0 license.
89.6KUpdated 4 hours agoMIT
macOS · Windows · Linux · Docker#Agent Client Protocol#Git integration#Human approval
OpenHands is a coding agent platform for developers and teams that want agents to make changes across a codebase and handle recurring engineering work. You can run it on your own computer or self-host it on a server. Its core is open source under the MIT license, and a separate OpenHands Cloud service offers hosted execution.
27.1KUpdated 3 hours ago
#Agent Skills#Streaming inference#Structured output
Vercel AI SDK is a TypeScript library for developers building chatbots, generative interfaces, and AI agents. It gives an application one way to work with models from providers such as OpenAI, Anthropic, and Google, so teams can change providers without rebuilding the parts of their app that handle responses. It suits developers who want control over the application they build while choosing how it connects to model services.
20.4KUpdated 2 months agoApache-2.0
macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration
gpt-oss is a pair of OpenAI reasoning models for developers who want to run a local LLM or host one on their own server. The models are open weight and licensed under Apache 2.0. OpenAI also has a hosted browser demo, separate from running the models on your hardware.
24.8KUpdated 22 hours agoApache-2.0
iOS · Android#Agent Skills#Hugging Face integration#Multilingual
Google AI Edge Gallery is an open-source app for people who want to try generative AI on their own phone. It runs model inference locally, so offline chat, image analysis and audio tasks don't send your inputs to a server. It supports Android and iOS and uses the Apache 2.0 license.
46Updated 5 days agoAGPL-3.0
macOS · Windows · Linux · Browser Extension#Code execution#GGUF#llama.cpp backend
Vyact is an open source desktop workspace for people who want to use a local LLM with their documents, email and code. It runs on Apple Silicon Macs, Windows and Linux x64 under the AGPL-3.0 license. Intel Macs aren't supported.
922Updated 5 months agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#LM Studio integration#MLX
Eclaire is a self-hosted AI assistant for notes, documents, photos, bookmarks and tasks. You can ask questions about saved material, inspect the sources behind an answer and ask the assistant to create notes or update tasks. Scheduled automations handle recurring requests such as a weekly task summary.
470Updated 1 month agoMIT
macOS#MCP#Tool calling
ast-grep MCP is a locally run server that lets AI coding assistants search code by its syntax structure. It connects ast-grep to Cursor, Claude Desktop and other clients that support the Model Context Protocol (MCP). It's for developers who need an assistant to find specific code constructs when matching words alone is too broad.
886Updated 2 weeks agoMIT
macOS · Windows · Linux · Docker · Web#Agent Skills#Code execution#Guardrails
PocketPaw is a self-hosted AI agent for people who want a personal assistant that can work with files, write code and use a browser on their own computer or server. It runs on macOS, Windows and Linux, with a web dashboard and access through messaging apps. It is open source under the MIT license.
42Updated 14 hours ago
macOS · Windows · Linux · Docker#Agent Skills#Code execution#Distributed execution
Code Buddy is an AI coding assistant for developers who want a terminal agent running on their own machine, with a choice of local or cloud models. It reads repositories, edits code and runs commands on Linux, macOS and Windows. Ollama keeps model inference local without an API key or account; cloud providers send model requests off the machine. Routing includes automatic provider failover.