Paperless-GPT adds AI text extraction and document organization to an existing paperless-ngx library. It runs in Docker on your own server and suits people who want less manual sorting of scanned paperwork. The project is open source under the MIT license.
It can generate document titles, tags, creation dates and correspondents, plus extract information into custom fields. Its web interface lets you review and edit suggestions before applying them, or let automatic processing handle documents. You can also adjust prompts and ask for summaries or specific information across a selection of documents.
For local LLM processing, it connects to Ollama, with support for reasoning models such as qwen3:8b and vision models such as minicpm-v. Using local backends keeps that processing on your hardware. Cloud alternatives include OpenAI, Anthropic Claude and Mistral, which receive the document text or images needed for the selected task and require API credentials.
OCR can use vision models or dedicated services, including self-hosted Docling Server, Google Document AI, Azure Document Intelligence and Mistral OCR. With Google Document AI, it can produce PDFs whose text is searchable and selectable while retaining the scanned page's appearance. It can save those files locally or upload them to paperless-ngx.
The web interface has no built-in authentication; access protection needs to come from your hosting setup.
Claim this page and we'll verify you by hand. Paperless-GPT gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Paperless-GPT?Promote it
Something wrong or outdated on this page?
68.2KUpdated 1 day agoMIT
macOS · Windows · Linux#MCP#Works offline
Docling is an MIT-licensed, open source document parser for developers turning files into structured content for search and AI applications. It runs locally on macOS, Linux, and Windows, including in air-gapped environments. Its PDF processing identifies page layout and reading order, extracts tables, code, and formulas, and classifies images.
37.3KUpdated 24 hours agoAGPL-3.0
macOS · Windows · Docker · Web#Batch processing#Hugging Face integration#MCP
9.4KUpdated 2 days agoMIT
macOS · Windows · Linux · Android · Docker · Web#Batch processing#LM Studio integration#MCP
20.3KUpdated 4 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning
29.4KUpdated 4 days agoAGPL-3.0
iOS · Android · Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
6.7KUpdated 2 months agoApache-2.0
Windows · Docker · Web#Batch processing#Hugging Face integration#Multilingual
PDFMathTranslate translates scientific PDFs while keeping their page layout, formulas, charts, contents pages and annotations. It's for researchers, students and others who need to read papers in another language without losing the relationship between the text and its figures. It produces both translated PDFs and bilingual documents for comparison with the original.
xberg, formerly Kreuzberg, is a local document extraction engine for developers building AI search, document processing, and retrieval-augmented generation applications. It reads PDFs, Office files, scanned images, email, and nested archives, extracting text, tables, images, and metadata through one shared engine. It's open source under MIT.
ebook2audiobook turns non-DRM ebooks into narrated audio with chapters and metadata, for readers who want audio editions of their own books. It runs locally on Windows, macOS and Linux, with Docker support and a browser interface built with Gradio. It's open source under Apache 2.0.
Karakeep is a bookmark manager for people who collect links, notes, images and PDFs in one place. You can run the open source app on your own server with Docker under the AGPL-3.0 license, or use the managed Karakeep Cloud service. It has a web app, iOS and Android apps, and extensions for Chrome, Firefox and Safari.
MonkeyOCR is a local AI document parser for developers and researchers working with English and Chinese PDFs or images. It extracts text, formulas and tables while identifying page structure and relationships between blocks. That makes it useful for documents where plain text extraction loses reading order or separates content from its layout.