Tools tagged with "OpenAI-compatible API"

100+ tools
Favicon of smolagents

smolagents

1 video
An open-source Python AI agent library with local models through Transformers or Ollama, cloud API support, and Docker sandboxing.

29.6KUpdated 1 week agoApache-2.0

#Code execution#Hugging Face integration#MCP

Self-hosted AI inference operator for Kubernetes with vLLM, Ollama and an OpenAI-compatible API. Runs on CPUs, GPUs or TPUs under Apache 2.0.

1.3KUpdated 2 days agoApache-2.0

Web#LoRA#Multimodal input#Ollama integration

An open-source Ruby AI framework for Ruby and Rails apps, with Ollama, hosted providers and OpenAI-compatible endpoints under an MIT license.

4.4KUpdated 1 day agoMIT

#Human approval#Multi-agent workflows#Multimodal input

An open-source OCR toolkit that converts PDFs and images into Markdown using a local GPU or an OpenAI-compatible inference server. Apache 2.0 licensed.

19.7KUpdated 6 months agoApache-2.0

Linux · Docker · Web#Batch processing#Distributed execution#OpenAI-compatible API

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

A self-hosted AI coding assistant for full-stack web apps, with Ollama and LM Studio support, desktop apps, and MIT-licensed source code.

19.9KUpdated 8 months agoMIT

macOS · Windows · Linux · Docker · Web#Code execution#Git integration#LM Studio integration

An open-source LLM client that brings local models through Ollama and llama.cpp, plus cloud services such as Claude and Gemini, into Emacs.

3.5KUpdated 6 days agoGPL-3.0

#Git integration#Human approval#llama.cpp backend

A self-hosted inference framework that coordinates NVIDIA GPU clusters with vLLM, SGLang or TensorRT-LLM and exposes an OpenAI-compatible API.

8.2KUpdated 21 hours ago

#Distributed execution#Multimodal input#OpenAI-compatible API

A local AI workbench for macOS, Windows and Linux. Evaluate agents, optimize prompts and run fully offline with Ollama or use cloud APIs.

5.1KUpdated 22 hours ago

macOS · Windows · Linux#Git integration#MCP#Multi-agent workflows

Favicon of MLflow

MLflow

4 videos
Self-hosted AI engineering platform for tracing agents, evaluating LLMs, and tracking models. Apache 2.0 licensed, with Ollama support.

28.2KUpdated 1 day agoApache-2.0

Docker · Web#Batch processing#LLM tracing#MCP

Local OCR software for PDFs and images, with reading order, tables and math. Runs on CPU, Apple Silicon or NVIDIA GPUs; code uses Apache 2.0.

21.4KUpdated 3 weeks agoApache-2.0

macOS · Web#Batch processing#llama.cpp backend#Multilingual

Favicon of KServe

KServe

2 videos
Self-hosted AI model serving platform for Kubernetes. Serve LLMs and predictive models with vLLM, Hugging Face support and an OpenAI-compatible API.

6KUpdated 23 hours agoApache-2.0

#Hugging Face integration#ONNX#OpenAI-compatible API

Favicon of WebLLM

WebLLM

2 videos
A local LLM engine that runs models in the browser with WebGPU acceleration, OpenAI API compatibility, and an Apache 2.0 license.

19.2KUpdated 2 weeks agoApache-2.0

Web · Browser Extension#OpenAI-compatible API#Streaming inference#Structured output

An open-source Python AI framework for self-hosted agents and document search, with local model support and an Apache 2.0 license.

26.6KUpdated 1 day agoApache-2.0

Docker#Guardrails#Hugging Face integration#Hybrid search

Self-hosted embedding and reranking API with MIT licensing, Hugging Face models, and CPU, NVIDIA, AMD and Apple MPS support.

2.9KUpdated 6 months agoMIT

macOS · Docker#Batch processing#Hugging Face integration#Multimodal input

A self-hosted AI answer engine that cites web sources and works with Ollama, local model servers, or cloud models.

36.9KUpdated 4 weeks agoMIT

Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API

Run AI models through Docker Desktop, Docker Engine or a standalone binary, with local inference and OpenAI and Ollama compatible APIs.

655Updated 2 days agoApache-2.0

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

An offline screen translator for Windows, macOS and Linux that reads Japanese game text and displays English overlays. Open source under MIT.

796Updated 2 weeks agoMIT

macOS · Windows · Linux#llama.cpp backend#LM Studio integration#Multilingual

A local AI desktop workspace for Apple Silicon Macs, Windows and Linux x64, with GGUF and MLX models, document search and optional cloud providers.

46Updated 5 days agoAGPL-3.0

macOS · Windows · Linux · Browser Extension#Code execution#GGUF#llama.cpp backend

Search your documents, organize notes and schedule tasks with a self-hosted AI assistant that connects to local text and vision models.

922Updated 5 months agoMIT

macOS · Windows · Linux · Docker · Web#llama.cpp backend#LM Studio integration#MLX