Tools tagged with "OpenAI-compatible API"

100+ tools
Favicon of Slotstream

Slotstream

1 video
Local LLM runner for Apple Silicon Macs that runs Qwen3.8-Flash-Next from SSD. Works offline after download and connects to coding agents and chat apps.

407Updated 2 days agoMIT

macOS#MLX#Multimodal input#OpenAI-compatible API

Favicon of PenEcho

PenEcho

1 video
An open-source AI canvas for macOS, Windows or a local browser, with MCP connections to coding agents and support for your own model API.

2.4KUpdated 3 days agoAGPL-3.0

macOS · Windows · Web#Code execution#MCP#Multimodal input

Favicon of Reef

Reef

1 video
Self-hosted AI agent infrastructure that learns from feedback, trains weights with Slime and SGLang, or improves prompts and skills without local training GPUs.

7.4KUpdated 13 hours agoApache-2.0

Linux#Agent Skills#OpenAI-compatible API#Prompt versioning

A browser-based local LLM tool that pools laptop, desktop and phone GPUs for chat and coding. Open source under MIT, with no account required.

544Updated 14 hours agoMIT

macOS · iOS · Web#Code execution#Distributed execution#Hugging Face integration

Favicon of Meetily

Meetily

6 videos
A local AI meeting assistant for macOS and Windows that records and transcribes calls offline, with Ollama or your own API key for summaries.

31.3KUpdated 3 weeks agoMIT

macOS · Windows · Linux#Ollama integration#OpenAI-compatible API#Streaming inference

Favicon of Bifrost

Bifrost

3 videos
A self-hosted AI gateway with an OpenAI-compatible API for Ollama and cloud providers, Apache 2.0 licensing, model fallbacks, and usage monitoring.

8.5KUpdated 23 hours agoApache-2.0

Web#LLM tracing#MCP#Multimodal input

Favicon of SurfSense

SurfSense

1 video
An open-source NotebookLM alternative for Windows, macOS and Linux. Ask questions with citations and create reports or podcasts using local models.

16.3KUpdated 2 days ago

macOS · Windows · Linux#LM Studio integration#Ollama integration#OpenAI-compatible API

Favicon of Graphiti

Graphiti

1 video
A self-hosted Python framework for AI agent memory that tracks changing facts. Apache 2.0 licensed, with support for local LLMs and cloud APIs.

31.3KUpdated 1 day agoApache-2.0

Docker#Hybrid search#Knowledge graphs#llama.cpp backend

Favicon of exo

exo

1 video
An open-source local LLM runner for macOS and Linux that splits models across devices and works offline with downloaded models. Apache 2.0 licensed.

47.7KUpdated 1 month agoApache-2.0

macOS · Linux · Web#Distributed execution#Hugging Face integration#MLX

Favicon of Jan

Jan

3 videos
A free desktop AI chat app for Windows, macOS and Linux that runs models locally or connects to GPT, Claude and other cloud models.

44.7KUpdated 5 hours ago

macOS · Windows · Linux#Hugging Face integration#MCP#OpenAI-compatible API

Favicon of vLLM

vLLM

9 videos
An open source LLM serving engine that runs on your hardware, supports NVIDIA and AMD GPUs, and provides an OpenAI-compatible API.

93KUpdated 2 hours agoApache-2.0

macOS · Docker#Batch processing#Distributed execution#GGUF

Favicon of LibreChat

LibreChat

3 videos
Self-hosted AI chat platform for local models and cloud APIs, with agents, searchable conversations, and multi-user access. MIT licensed.

45.2KUpdated 5 hours agoMIT

Web#Agent Skills#Code execution#MCP

Favicon of Dify

Dify

2 videos
A source-available AI agent and workflow builder that runs on your server with Docker or through a hosted cloud service.

157.6KUpdated 3 hours ago

Docker · Web#Code execution#MCP#OpenAI-compatible API

A terminal AI coding assistant that runs on macOS, Linux, Windows and BSD, with support for Ollama, llama.cpp and cloud model APIs.

28.4KUpdated 1 day ago

macOS · Windows · Linux · Android#llama.cpp backend#LM Studio integration#MCP

Self-hosted AI research notebook with Ollama and LM Studio support, source-based chat, and podcast generation. Open source under the MIT license.

39.6KUpdated 3 weeks agoMIT

Docker · Web#LM Studio integration#Ollama integration#OpenAI-compatible API

Favicon of LM Studio

LM Studio

16 videos
Download local language models, chat with documents and connect apps to a local model API on macOS, Windows or Linux.

lmstudio.aiComputer and Browser Agents

macOS · Windows · Linux#llama.cpp backend#MCP#MLX

Favicon of llama.cpp

llama.cpp

11 videos
An open source local LLM engine for GGUF models, with CPU and GPU support, a built-in web UI, and an OpenAI-compatible server.

130KUpdated 1 hour agoMIT

Web#Code execution#GGUF#Hugging Face integration

Open-source LLM fine-tuning framework under Apache 2.0, with LoRA, QLoRA, multimodal training and inference through vLLM or SGLang.

75.2KUpdated 2 days agoApache-2.0

Web#LoRA#Multimodal input#OpenAI-compatible API

Favicon of Cline

Cline

2 videos
An AI coding agent for editors, desktop, and terminal, with Apache 2.0 open source core and support for local models through Ollama and LM Studio.

69.6KUpdated 2 hours agoApache-2.0

macOS · Windows · VS Code · JetBrains#Code execution#Git integration#Human approval

Favicon of LiteLLM

LiteLLM

4 videos
A self-hosted AI gateway and Python SDK with an OpenAI-compatible interface for cloud and local models, spend controls, and request routing.

59.9KUpdated 1 hour ago

#Guardrails#MCP#Multi-user access

Favicon of Open WebUI

Open WebUI

14 videos
A self-hosted AI chat interface for Ollama, OpenAI-compatible APIs, and cloud models, with offline use and team access controls.

153.6KUpdated 1 week ago

Docker · Web#Code execution#Human approval#Hybrid search

Favicon of Ollama

Ollama

31 videos
Open-source local LLM runner for macOS, Windows, Linux and Docker, with optional cloud models and coding agent integrations.

182KUpdated 17 hours agoMIT

macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#Multimodal input

Favicon of SGLang

SGLang

3 videos
An open-source inference framework for serving language and multimodal models on your own hardware, with an OpenAI-compatible API.

36.7KUpdated 2 hours agoApache-2.0

#Batch processing#Distributed execution#LoRA

A self-hosted AI runtime with an OpenAI-compatible API. It runs models on CPUs or GPUs and keeps inference on your own hardware.

49.3KUpdated 2 hours agoMIT

macOS · Linux · Docker · Web#Code execution#Human approval#llama.cpp backend

Favicon of Zed

Zed

2 videos
A native code editor for macOS, Linux, and Windows that combines parallel AI agents with shared editing, chat, and screen sharing.

91.1KUpdated 21 hours ago

macOS · Windows · Linux#Agent Client Protocol#llama.cpp backend#LM Studio integration

Favicon of Unsloth

Unsloth

6 videos
An open-source local LLM app for macOS, Windows and Linux. Run and train models, generate media, and connect coding agents to your hardware.

77KUpdated 22 hours agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image

Local AI chat app for macOS that runs GGUF models offline, searches files on-device, and also connects to Claude, ChatGPT, and OpenAI-compatible endpoints.

recurse.chatChat With Your Documents

macOS#GGUF#Hugging Face integration#OpenAI-compatible API

An open-source AI gateway for self-hosted API management, connecting Ollama and cloud providers with shared authentication, access controls and usage reports.

1.8KUpdated 3 weeks agoApache-2.0

Web#MCP#Multi-user access#Ollama integration

An open source Python framework for AI red teaming, with automated attacks, a local web interface, and support for cloud services and custom endpoints.

4.6KUpdated 19 hours agoMIT

Web#AI red teaming#OpenAI-compatible API

Self-hosted AI gateway with OpenAI and Anthropic API compatibility, Ollama and vLLM support, caching, failover, and per-team usage tracking. MIT licensed.

1.2KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Guardrails#llama.cpp backend#LLM tracing

Self-hosted AI gateway routes OpenAI-compatible requests to cloud providers or vLLM, with centralized credentials and an Apache 2.0 license.

2.2KUpdated 20 hours agoApache-2.0

Docker#LLM tracing#MCP#Multi-user access

Self-hosted AI gateway with MIT licensing, per-key cost and rate limits, and support for OpenAI, Anthropic and local models through vLLM.

1.2KUpdated 2 years agoMIT

Docker#Guardrails#Multi-user access#OpenAI-compatible API

A self-hosted LLM routing framework in Python that connects to Ollama and cloud models through LiteLLM. Open source under Apache 2.0.

5.6KUpdated 2 years agoApache-2.0

#Ollama integration#OpenAI-compatible API

Self-hosted LLM API gateway with an OpenAI-compatible API, Ollama support and access controls. Runs as a single executable or in Docker under MIT.

37KUpdated 2 years agoMIT

Linux · Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API

A Home Assistant AI agent that controls devices, creates automations, and retrieves history using OpenAI or compatible backends such as LocalAI.

1.4KUpdated 3 weeks ago

Web#Home Assistant integration#OpenAI-compatible API#Tool calling

Self-hosted LLM inference for Kubernetes with NVIDIA, AMD and Apple Silicon support, OpenAI-compatible APIs, and an Apache 2.0 license.

223Updated 17 hours agoApache-2.0

macOS · Linux#GGUF#Git integration#Guardrails