Marqo Open Source is a self-hosted search engine for developers building semantic document search, image search, or retrieval for AI applications. It handles embedding generation alongside storage and retrieval, so applications can submit documents without maintaining a separate embedding service. The open-source project is deprecated and no longer receives updates.
It runs in Docker under the Apache 2.0 license and supports CPU or GPU inference. The stated Docker requirements are at least 8 GB of memory and 50 GB of storage. Self-hosting puts inference and search storage on your own infrastructure; Marqo Cloud is a separate managed offering.
Marqo supports PyTorch and Hugging Face models, including CLIP for searching images with text or other images. It also accepts custom models and supports OpenAI embeddings. Local models run on your infrastructure, while the OpenAI option uses an external service.
Search can use semantic similarity or keywords. Weighted queries let applications favor some concepts and reduce matches to others, while metadata filters narrow results. Documents can hold text, images, and structured metadata together. Combined text and image fields let both contribute to a document's relevance score.
Integrations with Haystack and LangChain connect the search engine to question answering and retrieval pipelines. Griptape and Hamilton integrations support agent and LLM applications. For larger collections, Marqo supports horizontal index sharding and asynchronous uploads and searches.
Claim this page and we'll verify you by hand. Marqo Open Source gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Marqo Open Source?Promote it
Something wrong or outdated on this page?
13KUpdated 22 hours agoApache-2.0
Docker#Agent Skills#Hugging Face integration#Knowledge graphs
txtai is a Python framework for developers building search applications, chat with their data, and AI agents on their own hardware or servers. Its embeddings database combines sparse and dense vector search with graphs and relational data, so the same system can find related content and supply context to language models. It's open source under Apache 2.0.
26.6KUpdated 6 days agoGPL-3.0
macOS · Linux · Docker#Hybrid search#Multimodal input#RAG
2.9KUpdated 6 months agoMIT
macOS · Docker#Batch processing#Hugging Face integration#Multimodal input
34.9KUpdated 4 weeks agoApache-2.0
Docker · Web#Agent Skills#Hybrid search#RAG
16.9KUpdated 1 day ago
Docker#Hybrid search#RAG#Reranking
4.7KUpdated 1 week agoApache-2.0
macOS · Windows · Linux · Docker#Hybrid search#Reranking#Semantic search
Typesense combines typo-tolerant site search with vector and semantic search in a self-hosted engine. It's for developers building searchable apps, product catalogs or AI search over their own data. The C++ engine uses an in-memory architecture for low-latency results as users type.
Infinity Embeddings is a self-hosted server for developers building semantic search and retrieval-augmented generation applications. It runs embedding and reranking models on your own hardware, with support for image and audio search alongside text. It's open source under MIT.
Qdrant is a self-hosted vector database for developers building semantic search, retrieval-augmented generation (RAG), recommendations and AI agent memory. It stores embeddings alongside JSON metadata, so applications can find similar content while restricting results by attributes such as location, text or numeric ranges. The Rust engine is open source under Apache 2.0 and runs locally in Docker or on your own servers.
Weaviate is a self-hosted vector database for developers building search applications, RAG systems, recommendation engines, and chatbots. It stores data objects alongside their vector embeddings, so applications can search by meaning and filter results using structured data. You can run the database locally with Docker, deploy it on Kubernetes, or use the hosted Weaviate Cloud service.
Infinity is a self-hosted database for developers building search and retrieval-augmented generation (RAG) into LLM applications. It combines embedding search with full-text search and structured filters, so an application can retrieve relevant records through both meaning and exact terms.