RAG Frameworks, Parsers and Vector Databases

What you need to build retrieval-augmented generation (RAG) yourself: pipeline frameworks, document parsers, embedding servers and vector databases.

Subcategories

96 tools
A self-hosted search database for LLM apps that combines vector and full-text retrieval. Runs in Docker or embedded in Python under Apache 2.0.

4.7KUpdated 1 week agoApache-2.0

macOS · Windows · Linux · Docker#Hybrid search#Reranking#Semantic search

Open-source OCR library for Node.js and Python that converts documents to Markdown using cloud vision models. MIT licensed; page images leave your machine.

12.3KUpdated 1 year agoMIT

Linux#Multimodal input#Structured output

Favicon of SurrealDB

SurrealDB

1 video
A self-hosted multi-model database for AI context and applications, with graph and vector search, embedded deployment and a source-available license.

33.1KUpdated 4 weeks ago

Web#Hybrid search#Knowledge graphs#MCP

Local OCR model that converts images and PDFs to text or Markdown. Runs on NVIDIA GPUs with vLLM or Transformers under the MIT license.

23.9KUpdated 8 months agoMIT

Linux#Batch processing#Hugging Face integration#Multimodal input

A Python reranking library with a shared interface for local models and cloud APIs, CPU inference through FlashRank, and an Apache 2.0 license.

1.6KUpdated 9 months agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

A local PDF extraction toolkit in Python with OCR, layout detection and table recognition. Runs on CPU or GPU and uses the AGPL-3.0 license.

10KUpdated 2 years agoAGPL-3.0

#Hugging Face integration#Multilingual

Local OCR and document parsing software that converts PDFs and images to Markdown, recognizes tables and formulas, and runs on NVIDIA GPUs.

6.7KUpdated 2 months agoApache-2.0

Windows · Docker · Web#Batch processing#Hugging Face integration#Multilingual

An open-source PHP AI framework for Symfony and Laravel, with local model support through Ollama, LM Studio and OpenAI-compatible services.

1.7KUpdated 2 days agoMIT

#LM Studio integration#Ollama integration#OpenAI-compatible API

Self-hosted AI document chat and agents with Ollama and Xinference support. Runs offline with local models on Windows, macOS and Linux.

38.7KUpdated 11 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API

A self-hosted LLM API server that runs ExLlamaV3 models on your hardware, with OpenAI-compatible endpoints and an AGPL-3.0 license.

1.4KUpdated 2 days agoAGPL-3.0

Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

Open-source AI data processing framework for local machines and Ray clusters, with multimodal cleaning, deduplication and Apache 2.0 licensing.

7.1KUpdated 2 days agoApache-2.0

Docker#Batch processing#Distributed execution#Multimodal input

Open-source Python toolkit for document layout detection and OCR workflows, with pretrained deep learning models and an Apache 2.0 license.

5.8KUpdated 4 years agoApache-2.0

An open-source document chat app that runs locally with Ollama. Query PDFs and other files through a browser interface. MIT licensed.

22.2KUpdated 1 month agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration

An open-source codebase digest tool for LLM prompts. Process local directories with Python or the CLI, self-host with Docker, or use the hosted website.

15.7KUpdated 1 year agoMIT

Docker · Web

Favicon of LangChain

LangChain

5 videos
An MIT-licensed Python framework for AI agents and LLM apps, with model, tool and data integrations. LangChain.js serves JavaScript and TypeScript.

147.3KUpdated 1 day agoMIT

#Human approval#RAG#Streaming inference

An open-source Java AI framework connects Spring applications to Ollama or cloud models, with document retrieval, tool calling and MCP support.

9.5KUpdated 21 hours agoApache-2.0

#MCP#Ollama integration#RAG

Open-source vector search library for C++ and Python, with MIT licensing, compressed indexes, and CPU or NVIDIA and AMD GPU support.

41KUpdated 23 hours agoMIT

#Batch processing#Semantic search

Favicon of Crawl4AI

Crawl4AI

1 video
Open-source web crawler that runs in Python or Docker, converts pages to Markdown and JSON, and offers a hosted API with MCP access.

84.5KUpdated 6 days agoApache-2.0

Docker#Structured output

Favicon of Chroma DB

Chroma DB

1 video
An open-source vector database for AI apps. Run it locally or self-host under Apache 2.0, or use managed search through Chroma Cloud.

29.4KUpdated 22 hours agoApache-2.0

#Semantic search

An open-source document chat app for Windows, macOS and Linux. Use local models through Ollama or llama-cpp-python, or connect cloud APIs.

25.8KUpdated 4 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access

An open-source search engine for typo-tolerant keyword and semantic search. Self-host on Linux, macOS or Docker, or use managed cloud hosting.

26.6KUpdated 6 days agoGPL-3.0

macOS · Linux · Docker#Hybrid search#Multimodal input#RAG

Favicon of LanceDB

LanceDB

1 video
An open source vector database that runs locally or in your cloud, with multimodal storage, hybrid search and Python, TypeScript and Rust SDKs.

11.6KUpdated 23 hours agoApache-2.0

#Hybrid search#Semantic search

Self-hosted AI document search and agents with source citations. Run local models through Ollama, vLLM or llama.cpp, including fully offline deployments.

18.3KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend

Open-source PostgreSQL extension for vector search with pgvector, DiskANN indexing and compression. Run it self-hosted or in Timescale Cloud.

3.1KUpdated 3 weeks agoPostgreSQL

macOS · Linux · Docker#Semantic search

A self-hosted text embedding server with a REST API, CPU and GPU support, and offline operation with downloaded model weights. Apache 2.0 licensed.

5.1KUpdated 1 week agoApache-2.0

macOS · Linux · Docker#Batch processing#Hugging Face integration#LLM tracing

An open-source local AI API under Apache 2.0 that connects to Ollama, llama.cpp and other OpenAI-compatible servers for document retrieval and agent workflows.

57.6KUpdated 1 week agoApache-2.0

Docker · Web#Code execution#llama.cpp backend#MCP

Open-source AI framework for semantic search, RAG and agents. Runs locally or in Docker, with Hugging Face, llama.cpp and cloud models via LiteLLM.

13KUpdated 22 hours agoApache-2.0

Docker#Agent Skills#Hugging Face integration#Knowledge graphs

Favicon of DSPy

DSPy

2 videos
Open-source Python framework for building modular LLM systems, with typed outputs, tool-using agents, and automatic prompt optimization. MIT licensed.

38.4KUpdated 4 days agoMIT

#Code execution#MCP#Multimodal input

An open-source Python RAG system under the MIT license that uses knowledge graphs and community summaries to answer questions about private datasets.

36.2KUpdated 7 days agoMIT

#Knowledge graphs#RAG#Semantic search

A self-hosted AI chatbot platform with a visual builder, local LLM support and human handoff. Runs on Docker or Kubernetes, with a hosted option.

322Updated 2 weeks agoMIT

iOS · Android · Docker · Web#Batch processing#Code execution#Guardrails

Self-hosted web extraction API converts pages and PDFs to Markdown for LLMs. Apache 2.0 service code runs in Docker; a hosted API is also available.

12.1KUpdated 4 months agoApache-2.0

Docker#Multimodal input#Structured output

Open source search and analytics suite built on Apache Lucene, with vector search for AI applications and Apache 2.0 licensing across its components.

13.8KUpdated 23 hours agoApache-2.0

#Semantic search

Self-hosted AI model serving platform for Linux, Windows and macOS. Run language, speech and image models through an OpenAI-compatible API under Apache 2.0.

9.6KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input

A self-hosted search engine that combines full-text and semantic search, stores vectors for RAG, and integrates with LangChain and MCP.

59.4KUpdated 1 day ago

#Hybrid search#MCP#Multilingual

Local document converter turns PDFs and Office files into Markdown, JSON or HTML, with OCR on CPU, NVIDIA GPUs or Apple Silicon and optional LLM support.

40.1KUpdated 3 weeks agoApache-2.0

macOS · Linux · Web#Batch processing#llama.cpp backend#Multilingual

A vector search SQLite extension written in C with no dependencies. Runs on Linux, macOS, Windows and in browsers via WebAssembly under Apache 2.0.

8.2KUpdated 4 months agoApache-2.0

macOS · Windows · Linux · Web#Semantic search