RAG Frameworks, Parsers and Vector Databases

What you need to build retrieval-augmented generation (RAG) yourself: pipeline frameworks, document parsers, embedding servers and vector databases.

Subcategories

96 tools
Favicon of LlamaIndex

LlamaIndex

2 videos
An MIT-licensed Python framework for document-based AI apps. It works with Ollama, while its separate document services run locally or in the cloud.

52.4KUpdated 2 days agoMIT

#Ollama integration#RAG#Reranking

A self-hosted AI agent builder with visual workflows and document retrieval. Runs through Docker, with hosted and commercial editions also available.

29.8KUpdated 1 day ago

Docker · Web#LLM tracing#MCP#Multi-user access

Favicon of Weaviate

Weaviate

1 video
A self-hosted vector database that combines semantic and keyword search with RAG. Run it locally with Docker, on Kubernetes, or in a managed cloud service.

16.9KUpdated 1 day ago

Docker#Hybrid search#RAG#Reranking

Favicon of Cognee

Cognee

2 videos
An open-source AI agent memory platform that runs locally on CPU, connects to Claude Code and Codex, and supports self-hosting or managed cloud hosting.

31.2KUpdated 22 hours agoApache-2.0

Docker · Web#Knowledge graphs#MCP#Multi-user access

Self-hosted AI inference operator for Kubernetes with vLLM, Ollama and an OpenAI-compatible API. Runs on CPUs, GPUs or TPUs under Apache 2.0.

1.3KUpdated 1 day agoApache-2.0

Web#LoRA#Multimodal input#Ollama integration

Open-source Java library for LLM apps on the JVM, with Ollama support, tool calling, RAG, and integrations for Spring Boot and Quarkus.

13.2KUpdated 1 day agoApache-2.0

#MCP#Ollama integration#RAG

An open-source Ruby AI framework for Ruby and Rails apps, with Ollama, hosted providers and OpenAI-compatible endpoints under an MIT license.

4.4KUpdated 1 day agoMIT

#Human approval#Multi-agent workflows#Multimodal input

An open-source OCR toolkit that converts PDFs and images into Markdown using a local GPU or an OpenAI-compatible inference server. Apache 2.0 licensed.

19.7KUpdated 6 months agoApache-2.0

Linux · Docker · Web#Batch processing#Distributed execution#OpenAI-compatible API

Self-hosted AI agent platform with document Q&A, workflows and MCP tools. Connect local DeepSeek, Llama or Qwen models, or use cloud providers.

22.9KUpdated 3 days agoGPL-3.0

Docker · Web#MCP#Multimodal input#RAG

A self-hosted OCR model that extracts text and page structure from PDFs and images, with vLLM, Hugging Face Transformers and CPU inference support.

9.2KUpdated 6 months agoMIT

Docker#Hugging Face integration#Multilingual#Multimodal input

Open-source Python toolkit for semantic search and RAG, with BGE embedding models, multilingual rerankers, evaluation and fine-tuning under MIT.

12.2KUpdated 1 month agoMIT

#Multilingual#Semantic search

A Python research assistant that answers questions with citations from local documents. Supports self-hosted models through LiteLLM; licensed under Apache 2.0.

9.3KUpdated 2 months agoApache-2.0

#Multilingual#Multimodal input#RAG

Favicon of Milvus

Milvus

2 videos
An open-source vector database for AI retrieval. Run it locally or on your servers, with hybrid search, metadata filtering and CPU or GPU acceleration.

46.3KUpdated 1 day agoApache-2.0

macOS · Linux#Hybrid search#Semantic search

An open-source Postgres search extension for full-text and vector retrieval, filters and aggregations. Run it locally or self-host it under AGPL-3.0.

9.3KUpdated 21 hours agoAGPL-3.0

Docker#MCP#Semantic search

A local AI workbench for macOS, Windows and Linux. Evaluate agents, optimize prompts and run fully offline with Ollama or use cloud APIs.

5.1KUpdated 22 hours ago

macOS · Windows · Linux#Git integration#MCP#Multi-agent workflows

Favicon of pgvector

pgvector

3 videos
A Postgres extension for self-hosted vector search on Linux, macOS and Windows, with exact matching and HNSW or IVFFlat indexes.

23.2KUpdated 1 day ago

macOS · Windows · Linux · Docker#Semantic search

A Python library for local embeddings and reranking, using ONNX Runtime with CPU or GPU support. Open source under Apache 2.0.

3.2KUpdated 24 hours agoApache-2.0

#Batch processing#Multilingual#ONNX

Local OCR software for PDFs and images, with reading order, tables and math. Runs on CPU, Apple Silicon or NVIDIA GPUs; code uses Apache 2.0.

21.4KUpdated 3 weeks agoApache-2.0

macOS · Web#Batch processing#llama.cpp backend#Multilingual

Favicon of PaddleOCR

PaddleOCR

2 videos
An open source OCR toolkit that runs locally, converts PDFs and images to Markdown or JSON, and supports multilingual text under Apache 2.0.

90.4KUpdated 2 weeks agoApache-2.0

Web#Multilingual#ONNX#Structured output

An open-source AI search platform for self-hosted retrieval and recommendations, with hybrid search, model inference and an Apache 2.0 license.

7.1KUpdated 1 day agoApache-2.0

Linux#Hybrid search#RAG#Semantic search

An open-source Python AI framework for self-hosted agents and document search, with local model support and an Apache 2.0 license.

26.6KUpdated 1 day agoApache-2.0

Docker#Guardrails#Hugging Face integration#Hybrid search

Python library for running and training embedding and reranker models locally, with Apache 2.0 licensing and pretrained models on Hugging Face.

19.1KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

Self-hosted embedding and reranking API with MIT licensing, Hugging Face models, and CPU, NVIDIA, AMD and Apple MPS support.

2.9KUpdated 6 months agoMIT

macOS · Docker#Batch processing#Hugging Face integration#Multimodal input

Favicon of Khoj

Khoj

2 videos
An open-source AI assistant for searching your documents and the web. Run it on your computer or use Khoj's cloud app with local or online LLMs.

37.5KUpdated 2 months agoAGPL-3.0

Web#RAG#Semantic search#Web search