
PasteGuard is a privacy proxy that runs locally or on your own server, including through Docker. It's for teams that use cloud AI but can't send raw customer records, client files or credentials to providers. It replaces sensitive values with placeholders before requests leave your environment and restores originals in supported responses.
The privacy layer runs on your hardware. In Mask Mode, the cloud provider still receives the surrounding prompt and its context, with detected private values replaced. Route Mode sends requests containing sensitive data to a local LLM through Ollama, vLLM or llama.cpp; requests without sensitive data can go to the configured cloud provider.
For apps and SDKs, PasteGuard works with OpenAI-compatible and Anthropic APIs. It also protects context sent by coding assistants such as Codex, Claude Code, Cursor and Windsurf, where logs, configuration files and project code can expose secrets.
Multilingual detection covers names, locations, emails, phone numbers and financial identifiers. Secret detection includes API keys, passwords, private keys, JWTs and connection strings. It combines format and checksum checks with GLiNER for names and places, and supports streaming responses.
A local dashboard records requests and shows what PasteGuard detected, masked and sent to the provider. The project is open source under Apache 2.0 and has no telemetry.
Claim this page with an email at pasteguard.com. PasteGuard gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find PasteGuard?Promote it
Something wrong or outdated on this page?
1.2KUpdated 2 days agoMIT
macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing
GoModel is a self-hosted AI gateway for developers and platform teams that want one API for local models and cloud providers. It accepts OpenAI- and Anthropic-compatible requests, so applications can keep their existing SDKs while the gateway handles provider selection and usage controls.
13.1KUpdated 4 months agoMIT
Docker · Web#Guardrails#Ollama integration#OpenAI-compatible API
9.5KUpdated 8 hours agoApache-2.0
Docker · Web#Guardrails#MCP#Tool calling
1.2KUpdated 2 years agoMIT
Docker#Guardrails#Multi-user access#OpenAI-compatible API
1.8KUpdated 3 weeks agoApache-2.0
Web#MCP#Multi-user access#Ollama integration
7.1KUpdated 3 days agoApache-2.0
Docker#Guardrails#LLM tracing#Multi-agent workflows
Portkey Gateway is a self-hosted AI gateway for developers whose apps need to use local models and cloud providers through one OpenAI-compatible API. It routes requests to Ollama, OpenAI, Anthropic, Google Gemini and other backends, with controls for handling failures and checking model inputs and outputs.
Higress is a self-hosted AI gateway for developers and teams managing model APIs and the tools their AI agents call. It puts LLM traffic and MCP servers behind a shared entry point, with authentication, traffic controls and monitoring. The open-source edition uses the Apache 2.0 license and runs locally in Docker without registration. Alibaba Cloud also offers a fully managed gateway.
BricksLLM is a self-hosted AI gateway for teams that need to control access and spending across LLM applications. It sits between applications and model providers, applying cost and rate limits to individual API keys. The gateway is open source under the MIT license and runs locally or on your server through Docker, with PostgreSQL and Redis.
APIPark is a self-hosted AI gateway and developer portal for teams that need to manage access to models and business APIs in one place. It connects Ollama alongside cloud providers such as OpenAI, Claude, Gemini and DeepSeek. The gateway runs on your infrastructure; requests to cloud providers still go to those services.
Plano, formerly Arch Gateway, is a self-hosted AI gateway for developers building applications with multiple agents or model providers. It puts routing, guardrails and request tracing in a separate service, so each agent doesn't need its own implementation of that infrastructure. It's open source under Apache 2.0.