
Bifrost is a self-hosted AI gateway for developers and teams whose applications use multiple model providers. It puts Ollama, custom model deployments, and cloud services behind one OpenAI-compatible API, so applications can switch models without maintaining a separate integration for each provider.
The gateway runs on infrastructure you control, either as an HTTP service with a web dashboard or embedded in a Go application. Model requests go to the selected backend: local models can use Ollama, while requests to OpenAI, Anthropic, AWS Bedrock, Google Vertex, and other cloud providers leave your infrastructure. Bifrost is open source under Apache 2.0, and commercial private deployments are also available.
Automatic fallbacks route requests to another model or provider when one fails. Load balancing distributes traffic across providers and API keys. Semantic caching reuses responses to similar requests to reduce repeated model calls, and the shared interface handles text, images, audio, and streaming.
Teams can manage spending and access through budgets and virtual keys. The MCP gateway centralizes connections to external tools, including databases and web search. A built-in dashboard and OpenTelemetry support help teams inspect usage and request behavior.
Bifrost works with existing OpenAI, Anthropic, Google GenAI, LiteLLM, and LangChain integrations. Its Go implementation emphasizes low request overhead, and custom plugins let teams add monitoring or application-specific logic.
Claim this page with an email at getmaxim.ai. Bifrost gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Bifrost?Promote it
Something wrong or outdated on this page?
9.5KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#MCP#Tool calling
Higress is a self-hosted AI gateway for developers and teams managing model APIs and the tools their AI agents call. It puts LLM traffic and MCP servers behind a shared entry point, with authentication, traffic controls and monitoring. The open-source edition uses the Apache 2.0 license and runs locally in Docker without registration. Alibaba Cloud also offers a fully managed gateway.
59.9KUpdated 55 minutes ago
#Guardrails#MCP#Multi-user access
42.4KUpdated 1 day agoApache-2.0
Docker · Web#Guardrails#Human approval#LLM tracing
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
11.7KUpdated 23 hours ago
Docker · Web#LLM tracing#MCP#Ollama integration
37.5KUpdated 4 days agoMIT
macOS · Windows · Linux · Docker · Web#LLM tracing#MCP#Multimodal input
LiteLLM gives platform teams one place to manage access to LLMs across providers. Its self-hosted AI gateway puts cloud services and internal or locally hosted models behind an OpenAI-compatible API, so applications can change models without changing their integration. Developers can also use its Python SDK directly.
Agno is a Python framework and runtime for developers building customer-facing or internal AI agents. You can run its agent platform locally with Docker, on your own servers or in your cloud. The open-source framework uses the Apache 2.0 license, and the platform keeps sessions, memory, knowledge and traces in your database.
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.
Arize Phoenix is a self-hosted platform for developers who need to understand why an AI agent failed and test changes before shipping them. It runs on a laptop, in Docker, or on Kubernetes. Self-hosting keeps traces on your infrastructure; Phoenix Cloud provides a hosted alternative. Phoenix uses the Elastic License 2.0 (ELv2), a source-available license.
Claude Code Router is an open-source local model gateway for developers who use coding agents and want to manage their model providers in one place. It runs on macOS, Windows and Linux, with Docker and a CLI with a browser interface also available. The project uses the MIT license.