LLMOps: Gateways, Observability and Guardrails

Self-hosted tools for LLM apps in production, including gateways that route between models, tracing, evals and guardrails.

Subcategories

98 tools
Favicon of Reef

Reef

1 video
Self-hosted AI agent infrastructure that learns from feedback, trains weights with Slime and SGLang, or improves prompts and skills without local training GPUs.

7.4KUpdated 12 hours agoApache-2.0

Linux#Agent Skills#OpenAI-compatible API#Prompt versioning

Favicon of Bifrost

Bifrost

3 videos
A self-hosted AI gateway with an OpenAI-compatible API for Ollama and cloud providers, Apache 2.0 licensing, model fallbacks, and usage monitoring.

8.5KUpdated 23 hours agoApache-2.0

Web#LLM tracing#MCP#Multimodal input

Favicon of agentacct

agentacct

1 video
Local coding agent dashboard tracks work, checks, tokens and estimated costs across Claude Code and Codex. MIT open source, with no account or cloud.

756Updated 5 hours agoMIT

macOS · Windows · Linux#LLM tracing#MCP

An open-source AI agent workspace for building and managing workflows, with Docker or Kubernetes self-hosting and centralized access and spend controls.

29.8KUpdated 1 day agoApache-2.0

macOS · Docker · Web#Code execution#LLM tracing#Multi-user access

Favicon of Mastra

Mastra

2 videos
TypeScript AI agent framework that runs locally or on your server, with memory, MCP, evaluations, and models from OpenAI, Anthropic, and Gemini.

28.4KUpdated 21 hours ago

Web#Human approval#LLM tracing#MCP

Favicon of Dify

Dify

2 videos
A source-available AI agent and workflow builder that runs on your server with Docker or through a hosted cloud service.

157.6KUpdated 3 hours ago

Docker · Web#Code execution#MCP#OpenAI-compatible API

Favicon of Opik

Opik

1 video
An open-source LLM observability platform for tracing and evaluating AI agents. Self-host it under Apache 2.0 or use Comet's hosted service.

22.3KUpdated 1 day agoApache-2.0

Web#Guardrails#LLM tracing#Prompt versioning

Favicon of Agno

Agno

1 video
A Python AI agent framework and runtime you can self-host with Docker, with database storage, a web control plane and an Apache 2.0 license.

42.4KUpdated 1 day agoApache-2.0

Docker · Web#Guardrails#Human approval#LLM tracing

Open-source computer-use agent tools with an MIT license, local VMs on Apple Silicon, and hosted desktop fleets across Linux, Windows, macOS, and Android.

27.3KUpdated 20 hours agoMIT

macOS · Windows · Linux#Code execution#MCP#Tool calling

Favicon of Langfuse

Langfuse

3 videos
Trace AI agents, evaluate outputs, and manage prompts with a self-hostable LLM observability platform whose core is MIT licensed.

35.2KUpdated 2 hours ago

Docker#Agent Skills#LLM tracing#MCP

Favicon of Pydantic AI

Pydantic AI

4 videos
Open source Python AI agent SDK with typed outputs, Ollama support, and an optional self-hosted model gateway. Licensed MIT.

20.3KUpdated 1 hour agoMIT

#Human approval#LLM tracing#MCP

Favicon of LiteLLM

LiteLLM

4 videos
A self-hosted AI gateway and Python SDK with an OpenAI-compatible interface for cloud and local models, spend controls, and request routing.

59.9KUpdated 57 minutes ago

#Guardrails#MCP#Multi-user access

Favicon of Promptfoo

Promptfoo

1 video
Open source CLI and library for local LLM evaluations and security testing, with Ollama and hosted model APIs plus CI/CD integration.

25.6KUpdated 1 hour agoMIT

#AI red teaming#Git integration#MCP

A self-hosted AI runtime with an OpenAI-compatible API. It runs models on CPUs or GPUs and keeps inference on your own hardware.

49.3KUpdated 2 hours agoMIT

macOS · Linux · Docker · Web#Code execution#Human approval#llama.cpp backend

A self-hosted LLMOps platform for prompt versioning, monitoring and caching, with Node.js, Python and LangChain clients. Apache 2.0 licensed.

3.3KUpdated 1 month agoApache-2.0

Docker · Web#Multi-user access#Prompt versioning

An open-source AI gateway for self-hosted API management, connecting Ollama and cloud providers with shared authentication, access controls and usage reports.

1.8KUpdated 3 weeks agoApache-2.0

Web#MCP#Multi-user access#Ollama integration

An open-source Python toolkit for checking LLM prompts and responses for injection, secrets and harmful content. MIT licensed; archived and unmaintained.

3.2KUpdated 3 months agoMIT

#Guardrails

An open source Python framework for AI red teaming, with automated attacks, a local web interface, and support for cloud services and custom endpoints.

4.6KUpdated 19 hours agoMIT

Web#AI red teaming#OpenAI-compatible API

An open-source prompt injection detector with a self-hosted server, OpenAI-based analysis and attack memory. Apache 2.0 licensed; archived.

1.5KUpdated 3 years agoApache-2.0

Web#Guardrails#Semantic search

Self-hosted LLM security scanner for prompts and responses. Use it as a Python library or REST API with local or OpenAI embeddings. Apache 2.0 licensed.

496Updated 3 years agoApache-2.0

Docker · Web#Guardrails#Semantic search

Self-hosted AI gateway with OpenAI and Anthropic API compatibility, Ollama and vLLM support, caching, failover, and per-team usage tracking. MIT licensed.

1.2KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing

Self-hosted AI gateway routes OpenAI-compatible requests to cloud providers or vLLM, with centralized credentials and an Apache 2.0 license.

2.2KUpdated 20 hours agoApache-2.0

Docker#LLM tracing#MCP#Multi-user access

Self-hosted AI gateway with MIT licensing, per-key cost and rate limits, and support for OpenAI, Anthropic and local models through vLLM.

1.2KUpdated 2 years agoMIT

Docker#Guardrails#Multi-user access#OpenAI-compatible API

A self-hosted LLM routing framework in Python that connects to Ollama and cloud models through LiteLLM. Open source under Apache 2.0.

5.6KUpdated 2 years agoApache-2.0

#Ollama integration#OpenAI-compatible API

Self-hosted LLM API gateway with an OpenAI-compatible API, Ollama support and access controls. Runs as a single executable or in Docker under MIT.

37KUpdated 2 years agoMIT

Linux · Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API

Self-hosted LLM inference for Kubernetes with NVIDIA, AMD and Apple Silicon support, OpenAI-compatible APIs, and an Apache 2.0 license.

223Updated 16 hours agoApache-2.0

macOS · Linux#GGUF#Git integration#Guardrails

Python toolkit for semantic search and RAG with BGE embedding models, multilingual rerankers, fine-tuning and evaluation. MIT licensed.

12.2KUpdated 1 month agoMIT

#Multilingual#Multimodal input#Semantic search

A coding assistant CLI that writes and runs code from plain-language requests. Runs locally or in Docker, with local models or OpenAI and Anthropic APIs.

55.1KUpdated 2 years agoMIT

Windows · Docker#Code execution#Multimodal input

AI content moderation models check text and images in LLM inputs and outputs. Use downloadable models or Meta's hosted Llama API.

4.4KUpdated 20 hours ago

#Guardrails#Multilingual#Multimodal input

An open source LLM programming language that combines Python logic with output constraints and supports local llama.cpp and Transformers models or cloud APIs.

4.2KUpdated 1 year agoApache-2.0

Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration

An open-source AI coding assistant for VS Code using Ollama, llama.cpp, LM Studio or hosted APIs, with an MIT-licensed self-hosted team gateway.

3.7KUpdated 1 day agoMIT

Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend

Discourse’s bundled AI plugin adds forum assistants, semantic search, summaries and moderation. Configure a self-hosted model service or cloud provider.

47.9KUpdated 1 day agoGPL-2.0

Linux · Web#LLM tracing#Multi-user access#Structured output

An open-source LLM fine-tuning tool with a browser interface. Runs on Ubuntu with NVIDIA GPUs or in Docker, under the Apache 2.0 license.

5.2KUpdated 4 days agoApache-2.0

Linux · Docker · Web#Hugging Face integration#LoRA#Quantization

A self-hosted LLM evaluation tool with a local dashboard, custom checks and root cause analysis. Apache 2.0 licensed; model grading can call cloud APIs.

2.4KUpdated 2 years agoApache-2.0

Docker · Web#Hugging Face integration#Ollama integration

LLM API benchmarking library in Python, licensed under Apache 2.0. Tests OpenAI-compatible endpoints and cloud providers. Archived and unmaintained.

1.1KUpdated 2 years agoApache-2.0

#OpenAI-compatible API

A desktop local LLM evaluation tool for macOS, Windows and Linux. Compare models through local or remote Ollama servers. Open source under MIT.

953Updated 3 weeks agoMIT

macOS · Windows · Linux#Ollama integration