Tools tagged with "Structured output"

59 tools
An open-source AI agent SDK for C#, Python and Java. Build workflows with local models through Ollama, LMStudio or ONNX, or connect to cloud providers.

28.6KUpdated 1 day agoMIT

macOS · Windows · Linux#LM Studio integration#MCP#Multi-agent workflows

Open-source browser agent SDK for local Chrome or Browserbase cloud browsers, with natural-language actions and structured extraction. MIT licensed.

25.5KUpdated 2 days agoMIT

#LLM tracing#Structured output

A Python library that constrains LLM output with grammars and JSON schemas, supporting local Transformers and llama.cpp backends plus OpenAI.

21.8KUpdated 4 months agoMIT

#llama.cpp backend#Structured output#Tool calling

A local LLM inference engine with OpenAI and Anthropic-compatible APIs. Runs on macOS, Linux and Windows with CPU, CUDA or Apple Silicon support.

7.7KUpdated 5 days agoMIT

macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Hugging Face integration

Favicon of DSPy

DSPy

2 videos
Open-source Python framework for building modular LLM systems, with typed outputs, tool-using agents, and automatic prompt optimization. MIT licensed.

38.4KUpdated 4 days agoMIT

#Code execution#MCP#Multimodal input

Open-source LLM serving toolkit for your own GPU servers, with quantization, text and vision models, and OpenAI-compatible APIs. Apache 2.0 licensed.

8.1KUpdated 3 days agoApache-2.0

#Batch processing#Distributed execution#Hugging Face integration

Favicon of FastMCP

FastMCP

2 videos
An open-source Python MCP framework for building servers, connecting to local or remote tools, and adding interactive interfaces to conversations. Apache 2.0.

27.9KUpdated 1 day agoApache-2.0

#MCP#Structured output#Tool calling

A self-hosted LLM inference library built on PyTorch for NVIDIA GPUs, with a Python API, OpenAI-compatible serving, and multi-node support.

14.7KUpdated 22 hours ago

Docker#Batch processing#Distributed execution#LoRA

Self-hosted web extraction API converts pages and PDFs to Markdown for LLMs. Apache 2.0 service code runs in Docker; a hosted API is also available.

12.1KUpdated 4 months agoApache-2.0

Docker#Multimodal input#Structured output

Favicon of CrewAI

CrewAI

3 videos
Open-source Python AI agent framework under MIT. Run workflows on your own infrastructure with Ollama models or connect to the OpenAI API.

59.2KUpdated 1 day agoMIT

#MCP#Multi-agent workflows#Ollama integration

An open-source Ruby AI framework for Ruby and Rails apps, with Ollama, hosted providers and OpenAI-compatible endpoints under an MIT license.

4.4KUpdated 1 day agoMIT

#Human approval#Multi-agent workflows#Multimodal input

.NET library for local LLM apps using Ollama, with streaming chat, embeddings and model management. Open source under MIT; also supports Ollama cloud.

1.4KUpdated 2 months agoMIT

#Multimodal input#Ollama integration#Streaming inference

An MIT-licensed LLM data extraction library that validates structured outputs with Pydantic and works with Ollama, llama-cpp-python, vLLM and cloud APIs.

14KUpdated 3 weeks agoMIT

#llama.cpp backend#Ollama integration#Streaming inference

Favicon of Outlines

Outlines

2 videos
An open-source Python library that constrains LLM outputs to schemas and grammars, with support for local backends including Ollama and llama.cpp.

15.9KUpdated 1 month agoApache-2.0

#llama.cpp backend#MLX#Ollama integration

Local LLM library runs GGUF models through llama.cpp in Node.js, Bun and Electron. MIT licensed, with GPU support and JSON schema enforcement.

2.2KUpdated 3 days agoMIT

macOS · Windows · Linux#Batch processing#GGUF#Guardrails

An open-source programming language for AI agents that runs on macOS, Linux and Windows, with typed model calls, local tracing and built-in evaluations.

9.4KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · iOS · Android · Web#Batch processing#LLM tracing#Structured output

A local AI workbench for macOS, Windows and Linux. Evaluate agents, optimize prompts and run fully offline with Ollama or use cloud APIs.

5.1KUpdated 22 hours ago

macOS · Windows · Linux#Git integration#MCP#Multi-agent workflows

Local OCR software for PDFs and images, with reading order, tables and math. Runs on CPU, Apple Silicon or NVIDIA GPUs; code uses Apache 2.0.

21.4KUpdated 3 weeks agoApache-2.0

macOS · Web#Batch processing#llama.cpp backend#Multilingual

Favicon of PaddleOCR

PaddleOCR

2 videos
An open source OCR toolkit that runs locally, converts PDFs and images to Markdown or JSON, and supports multilingual text under Apache 2.0.

90.4KUpdated 2 weeks agoApache-2.0

Web#Multilingual#ONNX#Structured output

Favicon of WebLLM

WebLLM

2 videos
A local LLM engine that runs models in the browser with WebGPU acceleration, OpenAI API compatibility, and an Apache 2.0 license.

19.2KUpdated 2 weeks agoApache-2.0

Web · Browser Extension#OpenAI-compatible API#Streaming inference#Structured output

Python client for Ollama with local and cloud model access, streaming chat, embeddings, and async support. Open source under the MIT license.

10.6KUpdated 2 days agoMIT

#Batch processing#Ollama integration#Streaming inference

A free TypeScript library for AI apps and agents, with streaming, provider switching, and chat UI hooks for React, Next.js, Svelte, and Vue.

27.1KUpdated 2 hours ago

#Agent Skills#Streaming inference#Structured output

Open-weight local LLMs under Apache 2.0, with 20B and 120B models that work with Ollama, LM Studio and vLLM.

20.4KUpdated 2 months agoApache-2.0

macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration