Tools tagged with "OpenAI-compatible API"

100+ tools
A free, MIT-licensed browser AI assistant that connects to Ollama, self-hosted models or cloud services for page summaries, translation and chat.

10.8KUpdated 2 days agoMIT

Browser Extension#Multilingual#Ollama integration#OpenAI-compatible API

Self-hosted AI chat and document Q&A runs on Linux, macOS and Windows with local or cloud models. Apache 2.0 licensed; archived and no longer maintained.

12KUpdated 12 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access

Local AI note-taking app for macOS, Windows and Linux, with Ollama support and note-based Q&A. Open source under AGPL-3.0; archived and unmaintained.

8.5KUpdated 1 year agoAGPL-3.0

macOS · Windows · Linux#Ollama integration#OpenAI-compatible API#RAG

Self-hosted document Q&A system with offline support, Chinese and English retrieval, and Ollama or OpenAI-compatible model connections. Licensed under AGPL-3.0.

14.2KUpdated 1 week agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API

A self-hosted text-to-speech server using Piper and Coqui XTTS v2, with voice cloning and an OpenAI-compatible API. Archived and no longer maintained.

857Updated 2 years agoAGPL-3.0

macOS · Windows · Linux · Docker#Multilingual#ONNX#OpenAI-compatible API

Self-hosted RAG chatbot with Ollama and cloud model support, licensed BSD-3-Clause. The project is archived and no longer maintained.

7.7KUpdated 4 months agoBSD-3-Clause

Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration

Deprecated Zep Community Edition provides self-hosted, time-aware agent memory using Docker and separate LLM and embedding services.

getzep/zepAgent Memory

Docker#Hybrid search#Knowledge graphs#OpenAI-compatible API

An open-source AI search engine you can self-host with Docker, with cited answers, Ollama support, and quick or multi-step web research.

9.2KUpdated 5 days agoApache-2.0

Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API

An open-source AI UI builder that runs locally with Docker, previews generated interfaces, and works with Ollama or cloud models.

22.6KUpdated 3 months agoApache-2.0

macOS · Docker · Web#Ollama integration#OpenAI-compatible API

A self-hosted AI coding assistant that builds apps with developer review. Runs locally, uses cloud model APIs, and is no longer maintained.

33.7KUpdated 4 months ago

Windows · Docker#Code execution#Human approval#Multi-agent workflows

An open-source Python library for preparing LLM training data locally or on Slurm and Ray clusters, with filtering, deduplication and generation.

3.4KUpdated 1 day agoApache-2.0

#Batch processing#Distributed execution#Multilingual

An open-source AI coding assistant for VS Code using Ollama, llama.cpp, LM Studio or hosted APIs, with an MIT-licensed self-hosted team gateway.

3.7KUpdated 1 day agoMIT

Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend

A self-hosted Matrix chatbot using OpenAI APIs, with encrypted-room support and conversation context. Archived and unmaintained.

241Updated 2 years agoAGPL-3.0

Docker#Multi-user access#OpenAI-compatible API

A local AI companion for Windows, macOS and Linux with voice chat, Live2D avatars and camera input. Runs offline with local models or connects to cloud APIs.

14KUpdated 5 months ago

macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP

An open source AI character interface you can run locally, with voice chat, VRM avatars, and support for Ollama, llama.cpp, and cloud APIs.

1.6KUpdated 1 year agoMIT

Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input

A browser-based LLM frontend for llama.cpp, koboldcpp, AI Horde and OpenAI-compatible APIs, with offline use and an AGPL-3.0 license.

752Updated 9 months agoAGPL-3.0

Web#llama.cpp backend#OpenAI-compatible API#Persistent memory

Local AI meeting notetaker for macOS, Windows and Linux. Keeps records on your device and supports local models or optional cloud AI. MIT-licensed.

9.4KUpdated 1 day agoMIT

macOS · Windows · Linux · iOS · Android#LM Studio integration#MCP#Ollama integration

LLM API benchmarking library in Python, licensed under Apache 2.0. Tests OpenAI-compatible endpoints and cloud providers. Archived and unmaintained.

1.1KUpdated 2 years agoApache-2.0

#OpenAI-compatible API

An open-source audio AI model for local transcription, translation and Q&A. Runs offline with vLLM or Transformers under Apache 2.0.

10.8KUpdated 3 months agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

A self-hosted document AI extension for Paperless-ngx that classifies files and searches your archive with Ollama or cloud APIs. MIT licensed.

6KUpdated 6 months agoMIT

Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API

Self-hosted LLM gateway with evaluation, A/B testing, and Ollama support. Open source under Apache 2.0; archived and no longer maintained.

11.7KUpdated 4 months agoApache-2.0

Docker · Web#Batch processing#LLM tracing#Multimodal input

A native macOS AI chat client that connects to Ollama, LM Studio and cloud APIs, with inline writing in other apps and locally stored data.

mindmac.appDesktop Chat Apps

macOS#llama.cpp backend#LM Studio integration#MLX

An archived AI CLI for shell pipelines on macOS, Linux and Windows. Connects to LocalAI or cloud APIs and saves conversations locally under an MIT license.

4.5KUpdated 7 months agoMIT

macOS · Windows · Linux#MCP#Ollama integration#OpenAI-compatible API

Self-hosted AI application server with an OpenAI-compatible API, local Ollama and vLLM backends, document search and agent tool calling. MIT licensed.

8.4KUpdated 21 hours agoMIT

#Agent Skills#Batch processing#Guardrails

A self-hosted ChatGPT alternative that runs Llama 2 and Code Llama locally, with an MIT license and an OpenAI-compatible API.

10.9KUpdated 3 years agoMIT

macOS · Docker · Web#GGUF#llama.cpp backend#OpenAI-compatible API

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An on-device AI SDK that runs text, image and audio models on macOS, Windows and Linux, with GGUF, MLX and an OpenAI-compatible API.

qualcomm/GenieXInference Libraries and Bindings

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

Self-hosted AI data assistant that queries databases, analyzes files, and generates reports with local models or cloud APIs. Open source under MIT.

20.1KUpdated 2 days agoMIT

macOS · Linux · Docker · Web#Code execution#llama.cpp backend#OpenAI-compatible API

An open source LLM evaluation framework for local models and hosted APIs, with GGUF, Hugging Face transformers and llama.cpp support. MIT licensed.

14.1KUpdated 2 weeks agoMIT

macOS#Batch processing#GGUF#Hugging Face integration

An open-source AI gateway built on Envoy. Self-host agent orchestration, route LLM requests, and capture OpenTelemetry traces under Apache 2.0.

7.1KUpdated 2 days agoApache-2.0

Docker#Guardrails#LLM tracing#Multi-agent workflows

An open-source AI gateway under the MIT license, with an OpenAI-compatible API for Ollama and cloud providers. Runs locally with Node.js or Docker.

13.1KUpdated 4 months agoMIT

Docker · Web#Guardrails#Ollama integration#OpenAI-compatible API

A self-hostable AI chat interface with Docker deployment, local model connections through LocalAI and RWKV-Runner, and desktop clients.

88.8KUpdated 2 months agoMIT

macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API

A local LLM family for chat, coding and multilingual tasks, with GGUF and Hugging Face formats for CPU or GPU use and support for llama.cpp and MLX.

127Updated 12 months ago

macOS#GGUF#Hugging Face integration#llama.cpp backend

A local image captioning model with open weights, Apache 2.0 code, SFW and NSFW coverage, and support for ComfyUI and vLLM.

1.3KUpdated 7 months agoApache-2.0

Windows · Docker#Hugging Face integration#Multimodal input#OpenAI-compatible API

A self-hosted LLM inference server under Apache 2.0, with Docker deployment, multi-GPU support and an OpenAI-compatible chat API. The project is archived.

10.9KUpdated 6 months agoApache-2.0

Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

Local AI voice assistant for Linux and Windows with camera vision, persistent memory and MCP tools. Uses Ollama or cloud APIs. MIT licensed.

5.7KUpdated 2 weeks agoMIT

macOS · Windows · Linux · Docker#Home Assistant integration#MCP#Multi-agent workflows