Tools tagged with "Reranking"

31 tools
Favicon of Graphiti

Graphiti

1 video
A self-hosted Python framework for AI agent memory that tracks changing facts. Apache 2.0 licensed, with support for local LLMs and cloud APIs.

31.3KUpdated 1 day agoApache-2.0

Docker#Hybrid search#Knowledge graphs#llama.cpp backend

Favicon of LightRAG

LightRAG

3 videos
An open-source RAG framework under the MIT license. Run it locally or in Docker with local models or hosted LLMs to query your documents.

39.9KUpdated 4 days agoMIT

macOS · Windows · Linux · Docker · Web#Knowledge graphs#LLM tracing#Multimodal input

Favicon of LibreChat

LibreChat

3 videos
Self-hosted AI chat platform for local models and cloud APIs, with agents, searchable conversations, and multi-user access. MIT licensed.

45.2KUpdated 5 hours agoMIT

Web#Agent Skills#Code execution#MCP

An open-source RAG engine for document search and AI agents, with self-hosted deployment, hybrid retrieval, citations and a separate cloud service.

91.5KUpdated 7 hours agoApache-2.0

macOS · Windows · Linux · Docker#Hybrid search#MCP#Multi-agent workflows

Favicon of llama.cpp

llama.cpp

11 videos
An open source local LLM engine for GGUF models, with CPU and GPU support, a built-in web UI, and an OpenAI-compatible server.

130KUpdated 1 hour agoMIT

Web#Code execution#GGUF#Hugging Face integration

Favicon of LiteLLM

LiteLLM

4 videos
A self-hosted AI gateway and Python SDK with an OpenAI-compatible interface for cloud and local models, spend controls, and request routing.

59.9KUpdated 1 hour ago

#Guardrails#MCP#Multi-user access

Favicon of Open WebUI

Open WebUI

14 videos
A self-hosted AI chat interface for Ollama, OpenAI-compatible APIs, and cloud models, with offline use and team access controls.

153.6KUpdated 1 week ago

Docker · Web#Code execution#Human approval#Hybrid search

An on-device document search engine that combines keyword and semantic search with local GGUF models, plus MCP access for AI agents. MIT licensed.

30.1KUpdated 3 weeks agoMIT

macOS#GGUF#Hugging Face integration#Hybrid search

Self-hosted query engine that joins live data sources and searches documents through SQL, with Docker deployment and MySQL or PostgreSQL clients.

24Updated 3 weeks ago

Docker · Web#Hybrid search#Reranking#Scheduled tasks

Self-hosted document Q&A system with offline support, Chinese and English retrieval, and Ollama or OpenAI-compatible model connections. Licensed under AGPL-3.0.

14.2KUpdated 1 week agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API

Deprecated Zep Community Edition provides self-hosted, time-aware agent memory using Docker and separate LLM and embedding services.

getzep/zepAgent Memory

Docker#Hybrid search#Knowledge graphs#OpenAI-compatible API

An open-source AI search engine you can self-host, with AGPL-3.0 licensing, live web research, clickable citations, and a choice of models.

11.9KUpdated 2 months agoAGPL-3.0

Web#Code execution#Multilingual#RAG

An open-source AI coding assistant for VS Code using Ollama, llama.cpp, LM Studio or hosted APIs, with an MIT-licensed self-hosted team gateway.

3.7KUpdated 1 day agoMIT

Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend

Python RAG framework for document-based AI apps, with local models through Ollama, cloud APIs, and customizable retrieval workflows.

39.6KUpdated 1 year ago

#Ollama integration#RAG#Reranking

Self-hosted RAG framework for document Q&A, with Docker, Ollama and Infinity support. Apache 2.0 licensed; archived and no longer maintained.

4.4KUpdated 7 months agoApache-2.0

Docker · Web#Batch processing#Multimodal input#Ollama integration

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

A self-hosted search database for LLM apps that combines vector and full-text retrieval. Runs in Docker or embedded in Python under Apache 2.0.

4.7KUpdated 1 week agoApache-2.0

macOS · Windows · Linux · Docker#Hybrid search#Reranking#Semantic search

A Python reranking library with a shared interface for local models and cloud APIs, CPU inference through FlashRank, and an Apache 2.0 license.

1.6KUpdated 9 months agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

An open-source document chat app that runs locally with Ollama. Query PDFs and other files through a browser interface. MIT licensed.

22.2KUpdated 1 month agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration

An open-source document chat app for Windows, macOS and Linux. Use local models through Ollama or llama-cpp-python, or connect cloud APIs.

25.8KUpdated 4 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access

An open-source LLM CLI for macOS, Linux and Windows that connects to Ollama or cloud providers and supports document chat, shell commands and AI agents.

10.5KUpdated 7 months agoApache-2.0

macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input

A self-hosted text embedding server with a REST API, CPU and GPU support, and offline operation with downloaded model weights. Apache 2.0 licensed.

5.1KUpdated 1 week agoApache-2.0

macOS · Linux · Docker#Batch processing#Hugging Face integration#LLM tracing

Favicon of llama-swap

llama-swap

1 video
A local AI proxy that switches models on demand through OpenAI and Anthropic compatible APIs. Runs on macOS, Windows, Linux and FreeBSD under MIT.

5.8KUpdated 2 days agoMIT

macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend

Favicon of LlamaIndex

LlamaIndex

2 videos
An MIT-licensed Python framework for document-based AI apps. It works with Ollama, while its separate document services run locally or in the cloud.

52.4KUpdated 2 days agoMIT

#Ollama integration#RAG#Reranking

Favicon of Weaviate

Weaviate

1 video
A self-hosted vector database that combines semantic and keyword search with RAG. Run it locally with Docker, on Kubernetes, or in a managed cloud service.

16.9KUpdated 1 day ago

Docker#Hybrid search#RAG#Reranking

Local LLM library runs GGUF models through llama.cpp in Node.js, Bun and Electron. MIT licensed, with GPU support and JSON schema enforcement.

2.2KUpdated 3 days agoMIT

macOS · Windows · Linux#Batch processing#GGUF#Guardrails

A Python research assistant that answers questions with citations from local documents. Supports self-hosted models through LiteLLM; licensed under Apache 2.0.

9.3KUpdated 2 months agoApache-2.0

#Multilingual#Multimodal input#RAG

A Python library for local embeddings and reranking, using ONNX Runtime with CPU or GPU support. Open source under Apache 2.0.

3.2KUpdated 24 hours agoApache-2.0

#Batch processing#Multilingual#ONNX

An open-source Python AI framework for self-hosted agents and document search, with local model support and an Apache 2.0 license.

26.6KUpdated 1 day agoApache-2.0

Docker#Guardrails#Hugging Face integration#Hybrid search

Python library for running and training embedding and reranker models locally, with Apache 2.0 licensing and pretrained models on Hugging Face.

19.1KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

Self-hosted embedding and reranking API with MIT licensing, Hugging Face models, and CPU, NVIDIA, AMD and Apple MPS support.

2.9KUpdated 6 months agoMIT

macOS · Docker#Batch processing#Hugging Face integration#Multimodal input