Tools tagged with "Structured output"

59 tools
Favicon of Graphiti

Graphiti

1 video
A self-hosted Python framework for AI agent memory that tracks changing facts. Apache 2.0 licensed, with support for local LLMs and cloud APIs.

31.3KUpdated 1 day agoApache-2.0

Docker#Hybrid search#Knowledge graphs#llama.cpp backend

Favicon of Firecrawl

Firecrawl

4 videos
An open source web scraping API you can self-host or use as a hosted service to turn websites into Markdown, JSON, HTML, or screenshots.

186.5KUpdated 1 day agoAGPL-3.0

Docker#Batch processing#MCP#Structured output

Favicon of vLLM

vLLM

9 videos
An open source LLM serving engine that runs on your hardware, supports NVIDIA and AMD GPUs, and provides an OpenAI-compatible API.

93KUpdated 2 hours agoApache-2.0

macOS · Docker#Batch processing#Distributed execution#GGUF

Favicon of Agno

Agno

1 video
A Python AI agent framework and runtime you can self-host with Docker, with database storage, a web control plane and an Apache 2.0 license.

42.4KUpdated 1 day agoApache-2.0

Docker · Web#Guardrails#Human approval#LLM tracing

Favicon of LM Studio

LM Studio

16 videos
Download local language models, chat with documents and connect apps to a local model API on macOS, Windows or Linux.

lmstudio.aiComputer and Browser Agents

macOS · Windows · Linux#llama.cpp backend#MCP#MLX

Favicon of llama.cpp

llama.cpp

11 videos
An open source local LLM engine for GGUF models, with CPU and GPU support, a built-in web UI, and an OpenAI-compatible server.

130KUpdated 1 hour agoMIT

Web#Code execution#GGUF#Hugging Face integration

Favicon of Pydantic AI

Pydantic AI

4 videos
Open source Python AI agent SDK with typed outputs, Ollama support, and an optional self-hosted model gateway. Licensed MIT.

20.3KUpdated 2 hours agoMIT

#Human approval#LLM tracing#MCP

Favicon of n8n

n8n

7 videos
A self-hosted workflow automation platform for building AI agents with a visual editor, custom code, and cloud or offline models.

206.4KUpdated 2 hours ago

#Code execution#Human approval#MCP

Favicon of Ollama

Ollama

31 videos
Open-source local LLM runner for macOS, Windows, Linux and Docker, with optional cloud models and coding agent integrations.

182KUpdated 17 hours agoMIT

macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#Multimodal input

Favicon of SGLang

SGLang

3 videos
An open-source inference framework for serving language and multimodal models on your own hardware, with an OpenAI-compatible API.

36.7KUpdated 2 hours agoApache-2.0

#Batch processing#Distributed execution#LoRA

An open source browser agent you can run locally from Python, with an MIT license and optional hosted agents and remote browsers.

116.8KUpdated 4 days agoMIT

#MCP#Ollama integration#Structured output

Self-hosted AI chat and document Q&A runs on Linux, macOS and Windows with local or cloud models. Apache 2.0 licensed; archived and no longer maintained.

12KUpdated 12 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access

Self-hosted AI research agent that supports Ollama, cloud models and MCP tools. MIT licensed; archived and no longer maintained.

12.7KUpdated 2 months agoMIT

Windows · Web#Human approval#MCP#Multi-agent workflows

An open source LLM programming language that combines Python logic with output constraints and supports local llama.cpp and Transformers models or cloud APIs.

4.2KUpdated 1 year agoApache-2.0

Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration

Discourse’s bundled AI plugin adds forum assistants, semantic search, summaries and moderation. Configure a self-hosted model service or cloud provider.

47.9KUpdated 1 day agoGPL-2.0

Linux · Web#LLM tracing#Multi-user access#Structured output

Self-hosted LLM gateway with evaluation, A/B testing, and Ollama support. Open source under Apache 2.0; archived and no longer maintained.

11.7KUpdated 4 months agoApache-2.0

Docker · Web#Batch processing#LLM tracing#Multimodal input

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

A local vision-language model for image analysis, document extraction and video understanding, with Apache 2.0 weights and Hugging Face Transformers support.

huggingface.coOCR and Document Scanning

Linux#Batch processing#Hugging Face integration#Multimodal input

A local LLM app that runs models offline on iOS and macOS, supports text and vision models, and uses ggml and llama.cpp under the MIT license.

2.1KUpdated 8 months agoMIT

macOS · iOS#llama.cpp backend#Multimodal input#RAG

A self-hosted LLM inference server under Apache 2.0, with Docker deployment, multi-GPU support and an OpenAI-compatible chat API. The project is archived.

10.9KUpdated 6 months agoApache-2.0

Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An open-source on-device AI framework for Android, iOS, desktop and web, with model conversion from PyTorch, TensorFlow and JAX.

3.5KUpdated 20 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Web#Agent Skills#Hugging Face integration#Multimodal input

An open-source document extraction engine with a Rust core, CPU-only processing, Docker deployment, and support for Ollama, LM Studio, and vLLM.

9.4KUpdated 2 days agoMIT

macOS · Windows · Linux · Android · Docker · Web#Batch processing#LM Studio integration#MCP

Open-source OCR library for Node.js and Python that converts documents to Markdown using cloud vision models. MIT licensed; page images leave your machine.

12.3KUpdated 1 year agoMIT

Linux#Multimodal input#Structured output

An open-source Python library for synthetic data and structured extraction, with Ollama, vLLM, cloud APIs and managed fine-tuning integrations.

1.7KUpdated 2 days agoApache-2.0

#Batch processing#Code execution#Multimodal input

An open-source React Native library that runs GGUF models on iOS and Android through llama.cpp, with GPU acceleration and image and audio understanding.

1KUpdated 3 days agoMIT

iOS · Android#GGUF#llama.cpp backend#Multilingual

A Python framework for generating and evaluating LLM datasets, with Apache 2.0 licensing and integrations for Anthropic, Cohere and Argilla.

3.4KUpdated 10 months agoApache-2.0

#Structured output

A self-hosted LLM API server that runs ExLlamaV3 models on your hardware, with OpenAI-compatible endpoints and an AGPL-3.0 license.

1.4KUpdated 2 days agoAGPL-3.0

Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An open-source Python AI agent framework under Apache 2.0 that connects agents across frameworks and supports tool use and human approval.

5KUpdated 23 hours agoApache-2.0

#Code execution#Human approval#MCP

Favicon of LangChain

LangChain

5 videos
An MIT-licensed Python framework for AI agents and LLM apps, with model, tool and data integrations. LangChain.js serves JavaScript and TypeScript.

147.3KUpdated 1 day agoMIT

#Human approval#RAG#Streaming inference

An open-source Java AI framework connects Spring applications to Ollama or cloud models, with document retrieval, tool calling and MCP support.

9.5KUpdated 21 hours agoApache-2.0

#MCP#Ollama integration#RAG

Favicon of Crawl4AI

Crawl4AI

1 video
Open-source web crawler that runs in Python or Docker, converts pages to Markdown and JSON, and offers a hosted API with MCP access.

84.5KUpdated 6 days agoApache-2.0

Docker#Structured output

An open-source AI email assistant for Thunderbird that connects to local models through Ollama or LM Studio, or to cloud services such as ChatGPT and Claude.

347Updated 2 days agoGPL-3.0

#Batch processing#LM Studio integration#Multilingual

JavaScript client for Ollama in Node.js and browsers. Connect to local models or Ollama's cloud with an open-source, MIT-licensed library.

4.4KUpdated 2 days agoMIT

Web#LoRA#Multimodal input#Ollama integration

An open-source AI browser automation tool with a no-code builder, a Playwright-compatible SDK, and support for Ollama and cloud models.

23.1KUpdated 1 day agoAGPL-3.0

#Code execution#MCP#Multi-agent workflows

An open-source vision model that runs locally with PyTorch and Hugging Face Transformers, supports CPU or CUDA GPUs, and uses the MIT license.

huggingface.coComputer Vision Models

#Hugging Face integration#Multimodal input#Structured output

Open-source Python framework for checking LLM inputs and outputs, generating structured data, and running guards in your app or a self-hosted service.

7.5KUpdated 1 month agoApache-2.0

#Guardrails#Structured output