← back to the directory
Model Retrieval & Memory

BGE

1BAAI's 2open 3text 4embedding 5model 6family

the six Ws · specification

W1 Who

The Beijing Academy of Artificial Intelligence (BAAI) develops the BGE embedding model family.

W2 What

BGE (BAAI General Embedding) is a family of open text embedding models, including bge-m3 and bge-large-en, used for dense retrieval and reranking.

W3 Where

Weights are hosted on Hugging Face under BAAI, with code in the FlagOpen/FlagEmbedding GitHub repository.

W4 When

First BGE models released in August 2023, with bge-m3 following in 2024.

W5 Why

Built to give the open-source community strong multilingual and long-document embeddings for RAG and search pipelines.

W6 With

Used via Sentence-Transformers, Hugging Face Transformers, or the FlagEmbedding library.

W7 Watch

Most BGE models are MIT licensed, with some variants under Apache 2.0 or Gemma license; fully open weights, no credentials required, self-hosted; maintained by BAAI.

BAAIEmbeddingRetrievalMITMultilingual

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.