← back to the directory
Model Retrieval & Memory

E5

1Microsoft's 2contrastively 3pretrained 4multilingual 5text 6embeddings

the six Ws · specification

W1 Who

Microsoft Research developed the E5 text embedding model family, maintained on Hugging Face under the intfloat account.

W2 What

E5 is a family of text embedding models, including e5-large-v2 and multilingual-e5-large, trained via weakly-supervised contrastive pretraining for retrieval and semantic search.

W3 Where

Weights are hosted on Hugging Face under the intfloat organization, with code in the microsoft/unilm GitHub repository.

W4 When

Original E5 released in December 2022, with multilingual-e5 variants following in 2023.

W5 Why

Built to provide strong general-purpose text embeddings competitive with proprietary APIs for search and RAG systems.

W6 With

Used via Sentence-Transformers or Hugging Face Transformers.

W7 Watch

MIT licensed, fully open weights, no credentials required, self-hosted; maintained by Microsoft (code) and the intfloat community account (weights).

MicrosoftEmbeddingRetrievalMITMultilingual

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.