Plano, formerly Arch Gateway, is a self-hosted AI gateway for developers building applications with multiple agents or model providers. It puts routing, guardrails and request tracing in a separate service, so each agent doesn't need its own implementation of that infrastructure. It's open source under Apache 2.0.
Built on Envoy and written primarily in Rust, Plano works with agents in any language or framework that expose an OpenAI-compatible chat completions API. Its orchestration uses purpose-built models such as Plano-Orchestrator to choose agents and coordinate their work within a conversation. Developers can add agents without rewriting application routing logic, or use the LLM gateway independently of agent orchestration.
Model requests can target a specific model, a named alias or an automatic choice based on preferences. Plano handles provider differences through a unified API, with OpenAI and Anthropic shown as supported providers. It also captures agent activity signals and OpenTelemetry traces, metrics and logs without requiring tracing code in each agent. Filter chains provide hooks for jailbreak protection, moderation policies and memory.
The proxy and Plano's routing models can run locally. A hosted service also runs Plano and its model family in the US-central region; requests using that service or external model providers go beyond your local deployment. Local hosting of the gateway doesn't make cloud model calls local.
Claim this page and we'll verify you by hand. Plano (formerly Arch Gateway) gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Plano (formerly Arch Gateway)?Promote it
Something wrong or outdated on this page?
42.4KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#Human approval#LLM tracing
Agno is a Python framework and runtime for developers building customer-facing or internal AI agents. You can run its agent platform locally with Docker, on your own servers or in your cloud. The open-source framework uses the Apache 2.0 license, and the platform keeps sessions, memory, knowledge and traces in your database.
1.2KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing
1.2KUpdated 2 years agoMIT
Docker#Guardrails#Multi-user access#OpenAI-compatible API
9.5KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#MCP#Tool calling
59.9KUpdated 55 minutes ago
#Guardrails#MCP#Multi-user access
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
GoModel is a self-hosted AI gateway for developers and platform teams that want one API for local models and cloud providers. It accepts OpenAI- and Anthropic-compatible requests, so applications can keep their existing SDKs while the gateway handles provider selection and usage controls.
BricksLLM is a self-hosted AI gateway for teams that need to control access and spending across LLM applications. It sits between applications and model providers, applying cost and rate limits to individual API keys. The gateway is open source under the MIT license and runs locally or on your server through Docker, with PostgreSQL and Redis.
Higress is a self-hosted AI gateway for developers and teams managing model APIs and the tools their AI agents call. It puts LLM traffic and MCP servers behind a shared entry point, with authentication, traffic controls and monitoring. The open-source edition uses the Apache 2.0 license and runs locally in Docker without registration. Alibaba Cloud also offers a fully managed gateway.
LiteLLM gives platform teams one place to manage access to LLMs across providers. Its self-hosted AI gateway puts cloud services and internal or locally hosted models behind an OpenAI-compatible API, so applications can change models without changing their integration. Developers can also use its Python SDK directly.
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.