
Docling is an MIT-licensed, open source document parser for developers turning files into structured content for search and AI applications. It runs locally on macOS, Linux, and Windows, including in air-gapped environments. Its PDF processing identifies page layout and reading order, extracts tables, code, and formulas, and classifies images.
It handles scanned files too. OCR reads scanned PDFs and images. Docling also parses DOCX, PPTX, XLSX, HTML, EPUB, Apple Pages, email, and OpenDocument files. Speech recognition supports audio files, while video parsing produces a transcript and representative keyframes. For bar, pie, and line charts, it can produce tables or code with descriptions.
Different file types feed into a shared DoclingDocument structure, which makes their contents easier to use in the same application. Docling can export Markdown, HTML, WebVTT, and lossless JSON, among other formats. It also handles specialized XML schemas for patents, journal articles, and XBRL financial reports, and supports visual language models such as GraniteDocling.
For retrieval and agent applications, Docling integrates with LangChain, LlamaIndex, Crew AI, and Haystack. An MCP server connects it to agents, and the separate docling-serve API server lets teams run it as a service. Developers can also use it through a Python library or command-line interface.
Claim this page and we'll verify you by hand. Docling gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Docling?Promote it
Something wrong or outdated on this page?
90.4KUpdated 2 weeks agoApache-2.0
Web#Multilingual#ONNX#Structured output
PaddleOCR is an open source OCR and document parsing toolkit for developers building document search, RAG systems and AI agents. It runs on your own hardware or a self-hosted server and turns PDFs and images into structured Markdown or JSON. The Python toolkit uses PaddlePaddle and carries the Apache 2.0 license.
80.8KUpdated 1 day ago
macOS · Windows · Linux#llama.cpp backend#MCP#MLX
MinerU parses documents locally into structured text for AI agents, RAG systems and knowledge bases. It's for people working with scanned PDFs, academic papers and Office files whose tables, formulas or page layouts need more care than plain text extraction.
9.4KUpdated 2 days agoMIT
macOS · Windows · Linux · Android · Docker · Web#Batch processing#LM Studio integration#MCP
xberg, formerly Kreuzberg, is a local document extraction engine for developers building AI search, document processing, and retrieval-augmented generation applications. It reads PDFs, Office files, scanned images, email, and nested archives, extracting text, tables, images, and metadata through one shared engine. It's open source under MIT.
23.9KUpdated 8 months agoMIT
Linux#Batch processing#Hugging Face integration#Multimodal input
DeepSeek-OCR is an open-source OCR model for developers building document processing tools and researchers studying how AI reads text through images. It runs on your own hardware with NVIDIA CUDA GPUs. Its distinctive focus is visual text compression: representing document images with compact sets of vision tokens for a language model to read.
76.8KUpdated 2 days agoApache-2.0
Windows#Multilingual
Tesseract is an open source OCR engine for extracting text from images, with a command line program and a library developers can embed in their own applications. It's suited to document processing workflows and software that needs text recognition. The project uses the Apache 2.0 license and doesn't include a graphical app.
15.5KUpdated 3 days agoApache-2.0
macOS · Windows · Linux · Docker#Multilingual
Unstructured is a local document processing library for developers building LLM applications and document ingestion pipelines. It turns PDFs, Word documents, HTML, emails and images into document elements that applications can use. The Python library is open source under Apache 2.0 and runs on your own hardware, including through Docker images for x86_64 and Apple Silicon.