
mxbai-embed is Mixedbread’s downloadable text embedding model family for developers building their own retrieval systems. It converts queries and passages into vectors that an application can compare for relevance or similarity.
The mxbai-embed-large-v1 model supports local inference with SentenceTransformers, Transformers and Transformers.js. Its official examples encode document batches and compare them with a query. Retrieval queries use the documented search prefix; documents do not need that prefix.
The model supports Matryoshka Representation Learning, letting developers shorten embeddings to reduce storage. The examples also show binary quantization of embedding vectors. These are changes to output representations, distinct from quantizing model weights.
The large-v1 weights are licensed under Apache-2.0 and target English text. The model card includes an Infinity Docker serving example and a hosted API option, but local inference does not require the hosted service. Mixedbread’s managed search platform and its newer platform-only models have separate capabilities and deployment terms.
Claim this page with an email at mixedbread.com. mxbai-embed gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find mxbai-embed?Promote it
Something wrong or outdated on this page?
2.8KUpdated 1 month agoMIT
macOS#Batch processing#Hugging Face integration#LoRA
ColPali is a local AI document retrieval library for developers and researchers building document search or retrieval-augmented generation systems. It searches pages as images, using their text, charts and layout together rather than relying on a separate OCR pipeline. The colpali-engine package is deprecated; its maintainers recommend Sentence Transformers for new projects and production use.
22.2KUpdated 1 week agoMIT
#Hugging Face integration#Multilingual
5.8KUpdated 2 days agoApache-2.0
#Hugging Face integration#Multilingual#Quantization
273Updated 2 years agoApache-2.0
Linux#Guardrails#Hugging Face integration#LM Studio integration
huggingface.coEmbedding and Reranker Models
#Batch processing#Hugging Face integration#Multilingual
2.2KUpdated 1 day agoMIT
#Hugging Face integration#Multilingual
Model2Vec turns sentence transformers into small static embedding models that run locally on CPU. It's for developers who need text embeddings for retrieval, code search or classification without the size and inference cost of the original transformer. The Python package is open source under the MIT license.
E5 Embeddings is a family of text embedding models for developers building search and retrieval systems on their own hardware. It converts text into numerical representations for matching queries with relevant passages. The family includes English and multilingual models, plus instruction-based variants for task-specific embeddings.
EmbeddingGemma is a text embedding model for developers building search and document features that run on phones, laptops or tablets. Based on Gemma 3, it converts text into numerical representations so applications can find related passages by meaning. Embeddings stay on your hardware, and the model works without an internet connection.
Granite is IBM's family of open-source AI models for developers and businesses that want to run and customize AI on their own hardware or servers. The language-model repository listed here is archived and no longer maintained. The broader family includes models for language, speech, document understanding and forecasting, released under Apache 2.0 for research and commercial use.
GTE (General Text Embedding) is Alibaba’s family of downloadable models for representing text as vectors. Developers use these representations to compare queries with documents, cluster related text or supply retrieval components for larger applications.