← back to the directory
Model Retrieval & Memory

Nomic Embed

1Nomic 2AI's 3open 4long-context 5text 6embedder

the six Ws · specification

W1 Who

Nomic AI develops the Nomic Embed text embedding model.

W2 What

Nomic Embed (nomic-embed-text-v1.5) is an open, 8192-token-context text embedding model that outperforms OpenAI's ada-002 on retrieval benchmarks.

W3 Where

Weights are on Hugging Face under nomic-ai, with training code on GitHub at nomic-ai/contrastors.

W4 When

Released in February 2024, with the Matryoshka-dimension v1.5 update following shortly after.

W5 Why

Built to provide a fully open, reproducible embedding model with auditable training data for retrieval and RAG applications.

W6 With

Used via Sentence-Transformers, Hugging Face Transformers, or the Nomic Atlas API.

W7 Watch

Apache 2.0 license, fully open weights and training data, no credentials required, self-hosted or via Nomic's hosted API; maintained by Nomic AI.

Nomic AIEmbeddingRetrievalApache-2.0Long-context

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.