
Outlines is an open-source Python library for developers who need LLM responses to match a defined structure. It constrains output during generation, reducing the need to repair malformed JSON or retry responses that don't fit an application's requirements. The library uses the Apache 2.0 license.
It works with local LLM backends including llama.cpp, Transformers and MLX-LM, as well as servers such as Ollama, vLLM, SGLang and TGI. It also connects to cloud providers including OpenAI, Anthropic and Gemini. Where inference runs depends on the backend you choose: local integrations run models on your hardware, while cloud integrations send requests to the selected provider. Dottxt offers a separate hosted API for structured generation without running your own models.
Output definitions can use Python types and Pydantic models, or express constraints through JSON Schema, regular expressions and context-free grammars. These cover tasks such as extracting a service ticket from a customer email, restricting a response to predefined choices, or generating an object with required fields. Outlines can also infer output structure from function signatures.
Its shared interface lets developers switch model providers without rewriting the surrounding generation code. It doesn't require an agent framework. For repeated requests, Outlines compiles constraints once rather than compiling them again for each generation.
Claim this page and we'll verify you by hand. Outlines gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Outlines?Promote it
Something wrong or outdated on this page?
77KUpdated 21 hours agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image
Unsloth brings model training and everyday AI use into a desktop app for people who want to run models on their own hardware. Its no-code interface covers chat, fine-tuning and media generation on macOS, Windows and Linux. The Unsloth software is open source under Apache 2.0.
31.3KUpdated 1 day agoApache-2.0
Docker#Hybrid search#Knowledge graphs#llama.cpp backend
14KUpdated 3 weeks agoMIT
#llama.cpp backend#Ollama integration#Streaming inference
1.7KUpdated 2 days agoApache-2.0
#Batch processing#Code execution#Multimodal input
21.8KUpdated 4 months agoMIT
#llama.cpp backend#Structured output#Tool calling
1KUpdated 3 days agoMIT
iOS · Android#GGUF#llama.cpp backend#Multilingual
Graphiti is a self-hosted Python framework for developers building AI agents that need to remember changing facts. It builds knowledge graphs from conversations, structured records and unstructured text, so an agent can query current information or recover what was true earlier. It's open source under Apache 2.0.
Instructor is an open-source library for developers who need structured data from local LLMs or cloud models. It turns natural-language input into typed objects that applications can use, with validation and retries built into the extraction process. Its focus is data extraction.
Curator is a Python library for developers preparing LLM training datasets or extracting structured records from existing data. It supports local inference through Ollama and vLLM alongside cloud model APIs, so the same data pipeline can use models on your hardware or a hosted provider. It's open source under Apache 2.0.
Guidance is an MIT-licensed, open source Python library for developers who need language model output to follow a defined format. It works with local LLM backends including Transformers and llama.cpp, as well as OpenAI's cloud service. It's for application code.
llama.rn brings llama.cpp into React Native apps so developers can run local LLM inference on iOS and Android. It's an MIT-licensed library for building AI features into a mobile app, with model processing on the device. It uses GGUF models and requires React Native's New Architecture.