MTEB (and MMTEB) what it measures, how it is scored and who leads it
Embeddings · MTEB community (Hugging Face and collaborators) · introduced Oct 13, 2022 · Mean task score
Scores from MTEB Multilingual v2, published Sep 22, 2026
About this benchmark
- Task
- Embed text for retrieval, clustering, classification, reranking and similarity.
- Dataset
- v1: 8 task types, 58 datasets, 112 languages. MMTEB: over 500 tasks in 250+ languages.
- Method
- Each embedding model is run on every task with the task's own metric; the leaderboard averages them.
- Metric
- Mean task score
- Organization
- MTEB community (Hugging Face and collaborators)
- Introduced
- Oct 13, 2022
- Official leaderboard
- huggingface.co/spaces/mteb/leaderboard
Versions
Newest first. New versions are added, never rewritten.
| Version | Date | |
|---|---|---|
| MMTEBMassive multilingual expansion (arXiv 2502.13595); our board reads MTEB Multilingual v2. | Feb 19, 20251 year ago | Massive multilingual expansion (arXiv 2502.13595); our board reads MTEB Multilingual v2. |
| 1Released with the paper. | Oct 13, 20223 years ago | Released with the paper. |
Sources: each benchmark's paper (arXiv) and its official site or leaderboard. Scores are the boards captured here every day, exactly as published.