4.4KUpdated 2 days agoMIT
Web#LoRA#Multimodal input#Ollama integration
Ollama JavaScript connects Node.js and browser applications to models running through Ollama. It's for developers building chat interfaces, AI agents or other apps that need a local LLM backend. The library is open source under the MIT license, with TypeScript types and an API that follows Ollama's REST interface.
15.8KUpdated 2 days agoApache-2.0
Web#Distributed execution#Hugging Face integration#LoRA
ms-swift is a Python framework for developers and researchers who want to train and deploy language or multimodal models on their own hardware. It brings fine-tuning, evaluation and model serving into one project, with support for Qwen3, DeepSeek-R1, Llama4 and Mistral, plus multimodal models such as Qwen3-VL and InternVL3.5. It's open source under Apache 2.0.
36.1KUpdated 2 months agoApache-2.0
VS Code · JetBrains#Code execution#Human approval#MCP
Continue is an open-source coding agent available as a CLI, VS Code extension and JetBrains plugin. Continue was acquired by Cursor, and the project is no longer actively maintained. Its repository is read-only, and its software uses the Apache 2.0 license.
28.2KUpdated 2 hours agoApache-2.0
macOS · Windows · Linux · Web · VS Code · JetBrains#Code execution#Git integration#MCP
Qwen Code is an Apache 2.0 licensed AI coding agent for developers who want help working through a codebase, changing code and checking the result. It runs on macOS, Windows and Linux, with a terminal interface and a desktop app. It builds on Google Gemini CLI and has developed into an agent that can use Qwen models alongside other model providers.
21.7KUpdated 22 hours agoApache-2.0
Docker · Web#Human approval#MCP#Multi-agent workflows
Google ADK is an open-source AI agent framework for developers building applications that carry out multi-step tasks. You can run agents locally in Docker or on your own infrastructure, and connect them to locally running models through adapters. The framework is optimized for Gemini but supports other models and providers. Its Python repository uses the Apache 2.0 license.
20.4KUpdated 3 months agoMIT
Web#Multimodal input#Tool calling
SWE-agent is an open-source AI coding agent that lets a language model use tools to attempt fixes for issues in GitHub repositories. It's aimed at developers and researchers studying how agents handle real software tasks. The project is in maintenance-only mode: mini-swe-agent has superseded it, and the maintainers recommend that successor for new users.
6.7KUpdated 10 months agoApache-2.0
#Tool calling
OLMo is a family of language models for researchers and developers who want to inspect, adapt or run a model on their own hardware. Weights are downloadable. Ai2 also provides the training data, code, checkpoints and reports behind the models, giving researchers material to study the full training process rather than only the finished model.
17.8KUpdated 1 day ago
Web#Git integration#Guardrails#LLM tracing
Wren AI is a self-hosted data agent for teams and agent builders who need answers based on agreed business definitions. It turns plain-language questions into SQL and interactive dashboards, using the same definitions when someone asks directly or through Claude, ChatGPT or Gemini over MCP.
10.6KUpdated 1 week agoMIT
macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend
llama-cpp-python brings llama.cpp model inference into Python applications and exposes it through a self-hosted OpenAI-compatible server. It's for developers building local AI applications or connecting existing API clients to models on their own hardware. The package is open source under the MIT license.
11.9KUpdated 4 days agoAGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
KoboldCpp pairs local model inference with a browser interface built for chat, creative writing and roleplay. A fork of llama.cpp, it bundles KoboldAI Lite with tools for keeping character details and story context alongside your conversations. It's open source under AGPL-3.0.
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
28.6KUpdated 1 day agoMIT
macOS · Windows · Linux#LM Studio integration#MCP#Multi-agent workflows
Semantic Kernel is an MIT-licensed SDK for developers adding AI agents to their applications. It supports local models through Ollama, LMStudio and ONNX, alongside cloud services such as OpenAI and Azure OpenAI. You choose the model backend. The SDK runs on Windows, macOS and Linux and supports C#, Python and Java.
23.7KUpdated 1 day agoApache-2.0
#Hugging Face integration#LoRA#Multimodal input
verl is a Python library for teams training large language models on their own GPU infrastructure. It's the open-source implementation of HybridFlow, aimed at researchers and engineers who need reinforcement learning after initial model training. It uses the Apache 2.0 license.
8.5KUpdated 2 days agoMIT
iOS · Android#GGUF#Hugging Face integration#llama.cpp backend
PocketPal AI is an open source assistant for people who want to run language models on a phone or tablet. It works on iOS, iPadOS and Android. Once you've downloaded a model, you can chat offline without an account, and your prompts, replies and documents stay on your device. The app is licensed under MIT.
2.9KUpdated 24 hours agoMIT
Web · VS Code#Code execution#Hugging Face integration#MCP
Inspect AI is a Python framework for researchers and developers testing language models and AI agents. Developed by the UK AI Security Institute and Meridian Labs, it evaluates coding, reasoning, knowledge, behavior and multimodal understanding, including tasks where agents must take actions to succeed.
21.8KUpdated 4 months agoMIT
#llama.cpp backend#Structured output#Tool calling
Guidance is an MIT-licensed, open source Python library for developers who need language model output to follow a defined format. It works with local LLM backends including Transformers and llama.cpp, as well as OpenAI's cloud service. It's for application code.
41.9KUpdated 6 days agoGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#MCP#Multimodal input#Ollama integration
Chatbox is an AI chat client for people who want local models and cloud providers in the same app. It connects to Ollama for local LLM use and supports GPT, Claude, Gemini, Grok and DeepSeek with your own API keys. Chatbox also offers its own hosted model service.
7.7KUpdated 5 days agoMIT
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Hugging Face integration
mistral.rs is an open source inference engine for running models on your own computer or self-hosted server. It's for developers building AI applications and people who want local chat, multimodal models and agent tools in the same runtime. The Rust project uses the MIT license.
18.3KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend
DocsGPT is an MIT-licensed, open-source platform for teams that want AI search, assistants and agents over their own documents. It can run on your servers with local models, including fully air-gapped deployments where documents and questions stay inside your network. Answers include the source title and page number so readers can check the evidence.
91Updated 1 day agoAGPL-3.0
Web#Multi-user access#Multilingual#Multimodal input
Nextcloud Assistant brings AI into Nextcloud Hub's documents, email, chat and calendar. It's for teams that want help with shared work while choosing where AI processing happens. With an on-premises model, data stays on your server. The app is open source under AGPL-3.0.
7.8KUpdated 2 days agoAGPL-3.0
#LM Studio integration#Multi-agent workflows#Ollama integration
Obsidian Copilot brings AI agents into your vault to find notes by meaning, edit drafts and organize research. It's for writers, researchers and other Obsidian users who want AI to work directly with their existing notes. Agents create Markdown files, wikilinks and canvases that remain in your vault.
21.7KUpdated 2 months agoApache-2.0
Web#Persistent memory#RAG#Tool calling
Coze Studio is a self-hosted AI agent development platform for developers and teams building assistants and AI apps through visual, low-code tools. It brings agent creation, debugging and publishing into one workspace, with workflows for defining business logic. The open-source edition uses the Apache 2.0 license and shares its core engine with the commercial Coze Development Platform.
57.6KUpdated 1 week agoApache-2.0
Docker · Web#Code execution#llama.cpp backend#MCP
PrivateGPT is a self-hosted API layer for developers building AI applications around local models. It adds document retrieval, database access and agent tools to an existing model server. Local workflows can work offline and keep data within your environment; web search and connections to online providers need internet access.
9.1KUpdated 23 hours agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access
Local Deep Research is a self-hosted AI research assistant for people who need cited answers drawn from academic papers, the web and their own documents. It can produce a quick summary or pursue a complex question through repeated searches, then assemble a structured report. It's open source under MIT.
13KUpdated 23 hours agoApache-2.0
Docker#Agent Skills#Hugging Face integration#Knowledge graphs
txtai is a Python framework for developers building search applications, chat with their data, and AI agents on their own hardware or servers. Its embeddings database combines sparse and dense vector search with graphs and relational data, so the same system can find related content and supply context to language models. It's open source under Apache 2.0.
38.4KUpdated 4 days agoMIT
#Code execution#MCP#Multimodal input
DSPy is a Python framework for developers building AI applications whose tasks need clear inputs, predictable output types, and measurable results. You define what a language model should produce, then compose those tasks into a larger program. It's open source under the MIT license.
8.1KUpdated 3 days agoApache-2.0
#Batch processing#Distributed execution#Hugging Face integration
LMDeploy is an open-source toolkit for developers serving language and vision-language models on their own hardware. It combines model compression with inference and self-hosted APIs, so teams can use it for batch processing or as the model backend for an application. It uses the Apache 2.0 license.
322Updated 2 weeks agoMIT
iOS · Android · Docker · Web#Batch processing#Code execution#Guardrails
Tiledesk is a self-hosted platform for building AI agents and connecting them to human support teams. It's aimed at businesses automating customer conversations, internal information searches and workflows through a visual builder. You can run it on your own server with Docker or Kubernetes, or use its hosted cloud service.
27.9KUpdated 1 day agoApache-2.0
#MCP#Structured output#Tool calling
FastMCP is an open-source Python framework for developers connecting AI agents to their own tools and data through the Model Context Protocol (MCP). It supports locally running servers and connections to remote servers, with server development and client access in the same framework. It uses the Apache 2.0 license.
37.7KUpdated 2 days agoApache-2.0
macOS · Windows · Linux · Docker#Code execution#LM Studio integration#MCP
Playwright MCP lets AI agents control a browser by reading structured page information rather than interpreting screenshots. It's for developers building browser automation, exploratory tests or agent workflows that need to keep a browser session open across repeated actions. The server runs locally on macOS, Windows and Linux, or as a self-hosted service.
9.6KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input
Xinference serves language, speech and multimodal models through a shared API on your own computer or servers. It's an open source platform under Apache 2.0 for developers and researchers who want to build applications around models they host. You can also deploy it on cloud infrastructure.
5.8KUpdated 2 days agoMIT
macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend
llama-swap is a self-hosted proxy for people running several AI models on their own hardware. It starts the model server a request needs and swaps out another when necessary, so you don't have to keep every model loaded or manage separate API connections in your apps.
29.6KUpdated 1 week agoApache-2.0
#Code execution#Hugging Face integration#MCP
smolagents is an open-source Python library for developers building AI agents that carry out tasks by writing and executing Python. Its CodeAgent can combine tool calls with loops, conditionals, and calculations in one action. This approach suits tasks that need several operations, rather than a single model response.
59.2KUpdated 1 day agoMIT
#MCP#Multi-agent workflows#Ollama integration
CrewAI is an open-source Python framework for developers building workflows with multiple AI agents on their own hardware or servers. It supports local models through Ollama and defaults to the OpenAI API for model requests. The framework uses the MIT license; a separate commercial platform provides managed deployment and governance, with on-premise and cloud options.
29.8KUpdated 1 day ago
Docker · Web#LLM tracing#MCP#Multi-user access
FastGPT is a self-hosted AI agent builder for teams that want assistants to answer questions using company documents and carry out business workflows. Its visual editor connects model calls, knowledge retrieval and tools into applications for customer support, internal knowledge search and document review. You can run the platform on your own server through Docker or use the vendor's hosted service.
26.8KUpdated 2 months agoApache-2.0
Web#Code execution#Git integration#Multimodal input
Onlook is a visual code editor for designers who want to work directly on a website rather than hand off a separate mockup. It uses the actual app code as the design source, so visual edits change the product itself. The Apache 2.0 project can run locally or be self-hosted; Onlook also offers a hosted cloud product.