
Kong Gateway puts API, LLM and MCP traffic behind a shared gateway on your own infrastructure. It's for platform teams that need consistent access controls and traffic policies across services and AI applications. The open-source gateway uses the Apache 2.0 license and runs natively on Kubernetes through Kong's official Ingress Controller.
Its core capabilities cover routing, load balancing, health checks and authentication. Teams can apply rate limits, handle TLS connections, and use plugins to modify requests and responses or add logging and monitoring. Plugins also let developers extend the gateway for their own services.
For AI applications, a common LLM API routes requests across providers including OpenAI, Anthropic, GCP Gemini, AWS Bedrock, Azure AI and Mistral. AI capabilities include semantic routing and caching, security checks and observability. Kong also governs MCP traffic, records analytics and can generate MCP interfaces from RESTful APIs.
The gateway can run on infrastructure you control, but requests to cloud model providers go to those services. Kong Konnect is a separate commercial cloud offering with a managed control plane, analytics, a service catalog and developer portals.
Deployment choices include running without a database and separating the control plane from the components that handle traffic. That separation lets teams manage gateway policies centrally while operating the traffic layer on their own infrastructure.
The Apache 2.0 license applies to the open-source gateway. Advanced AI and MCP plugins may require a commercial edition; check each plugin's deployment and license requirements.
Claim this page with an email at konghq.com. Kong Gateway gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Kong Gateway?Promote it
Something wrong or outdated on this page?
9.5KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#MCP#Tool calling
Higress is a self-hosted AI gateway for developers and teams managing model APIs and the tools their AI agents call. It puts LLM traffic and MCP servers behind a shared entry point, with authentication, traffic controls and monitoring. The open-source edition uses the Apache 2.0 license and runs locally in Docker without registration. Alibaba Cloud also offers a fully managed gateway.
2.2KUpdated 20 hours agoApache-2.0
Docker#LLM tracing#MCP#Multi-user access
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
1.2KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing
59.9KUpdated 57 minutes ago
#Guardrails#MCP#Multi-user access
1.2KUpdated 2 years agoMIT
Docker#Guardrails#Multi-user access#OpenAI-compatible API
Agent Router is an open source AI gateway for teams whose agents use both model APIs and MCP tools. It runs on a laptop, a dedicated gateway, or Kubernetes, and gives applications one OpenAI-compatible entry point for cloud providers and self-hosted inference. The project uses the Apache 2.0 license.
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.
GoModel is a self-hosted AI gateway for developers and platform teams that want one API for local models and cloud providers. It accepts OpenAI- and Anthropic-compatible requests, so applications can keep their existing SDKs while the gateway handles provider selection and usage controls.
LiteLLM gives platform teams one place to manage access to LLMs across providers. Its self-hosted AI gateway puts cloud services and internal or locally hosted models behind an OpenAI-compatible API, so applications can change models without changing their integration. Developers can also use its Python SDK directly.
BricksLLM is a self-hosted AI gateway for teams that need to control access and spending across LLM applications. It sits between applications and model providers, applying cost and rate limits to individual API keys. The gateway is open source under the MIT license and runs locally or on your server through Docker, with PostgreSQL and Redis.