Skip to content
@embeddings-benchmark

Massive Text Embedding Benchmark

MTEB is a Python framework for evaluating embeddings and retrieval systems for both text and image. MTEB covers more than 1000 languages and diverse tasks, from classics like classification and clustering to use-case specialized tasks such as legal, code, or healthcare retrieval.

You can get started using mteb.

Overview
📈 Leaderboard The interactive leaderboard of the benchmark
Get Started.
🏃 Get Started Overview of how to use mteb
🤖 Defining Models How to use existing model and define custom ones
📋 Selecting tasks How to select tasks, benchmarks, splits etc.
🏭 Running Evaluation How to run the evaluations, including cache management, speeding up evaluations etc.
📊 Loading Results How to load and work with existing model results
Overview.
📋 Tasks Overview of available tasks
📐 Benchmarks Overview of available benchmarks
🤖 Models Overview of available Models
Contributing
🤖 Adding a model How to submit a model to MTEB and to the leaderboard
👩‍💻 Adding a dataset How to add a new task/dataset to MTEB
👩‍💻 Adding a benchmark How to add a new benchmark to MTEB and to the leaderboard
🤝 Contributing How to contribute to MTEB and set it up for development

Popular repositories Loading

  1. mteb mteb Public

    MTEB: State-of-the-art evaluation of embeddings across languages and modalities

    Python 3.4k 702

  2. results results Public

    Data for the MTEB leaderboard

    Python 61 196

  3. embedders-dilemma embedders-dilemma Public

    Code, datasets, and results for "The Embedder's Dilemma: LLMs Are Better, but at What Cost?" (COLM 2026). MTEB(LLM) benchmarks 10 LLMs and 26 embedding models on 37 tasks with full cost and through…

    Python 34 1

  4. leaderboard leaderboard Public archive

    Code for the MTEB leaderboard

    Python 32 15

  5. arena arena Public

    Code for the MTEB Arena

    Python 25 12

  6. mtebpaper mtebpaper Public

    Resources & scripts for the paper "MTEB: Massive Text Embedding Benchmark"

    Python 18 5

Repositories

Showing 10 of 11 repositories
  • MTEB-gym-v2 Public

    Offline LLM-judged arena for embedding models

    embeddings-benchmark/MTEB-gym-v2's past year of commit activity
    Python 2 3 17 5 Updated Sep 16, 2026
  • mteb Public

    MTEB: State-of-the-art evaluation of embeddings across languages and modalities

    embeddings-benchmark/mteb's past year of commit activity
    Python 3,425 Apache-2.0 702 273 (3 issues need help) 48 Updated Sep 16, 2026
  • results Public

    Data for the MTEB leaderboard

    embeddings-benchmark/results's past year of commit activity
    Python 61 CC0-1.0 196 0 6 Updated Sep 16, 2026
  • embeddings-benchmark/leaderboard-frontend's past year of commit activity
    Svelte 2 Apache-2.0 5 1 3 Updated Aug 20, 2026
  • embedders-dilemma Public

    Code, datasets, and results for "The Embedder's Dilemma: LLMs Are Better, but at What Cost?" (COLM 2026). MTEB(LLM) benchmarks 10 LLMs and 26 embedding models on 37 tasks with full cost and throughput accounting.

    embeddings-benchmark/embedders-dilemma's past year of commit activity
    Python 34 Apache-2.0 1 0 0 Updated Aug 19, 2026
  • .github Public
    embeddings-benchmark/.github's past year of commit activity
    0 0 0 0 Updated Aug 18, 2026
  • autoembed Public
    embeddings-benchmark/autoembed's past year of commit activity
    Python 0 1 1 0 Updated Aug 1, 2026
  • arena Public

    Code for the MTEB Arena

    embeddings-benchmark/arena's past year of commit activity
    Python 25 12 25 5 Updated Jul 2, 2025
  • miebpaper Public
    embeddings-benchmark/miebpaper's past year of commit activity
    Jupyter Notebook 2 0 0 0 Updated Feb 28, 2025
  • leaderboard Public archive

    Code for the MTEB leaderboard

    embeddings-benchmark/leaderboard's past year of commit activity
    Python 32 15 15 2 Updated Feb 4, 2025