Skip to content
@embeddings-benchmark

Massive Text Embedding Benchmark

MTEB is a Python framework for evaluating embeddings and retrieval systems for both text and image. MTEB covers more than 1000 languages and diverse tasks, from classics like classification and clustering to use-case specialized tasks such as legal, code, or healthcare retrieval.

You can get started using mteb.

Overview
📈 Leaderboard The interactive leaderboard of the benchmark
Get Started.
🏃 Get Started Overview of how to use mteb
🤖 Defining Models How to use existing model and define custom ones
📋 Selecting tasks How to select tasks, benchmarks, splits etc.
🏭 Running Evaluation How to run the evaluations, including cache management, speeding up evaluations etc.
📊 Loading Results How to load and work with existing model results
Overview.
📋 Tasks Overview of available tasks
📐 Benchmarks Overview of available benchmarks
🤖 Models Overview of available Models
Contributing
🤖 Adding a model How to submit a model to MTEB and to the leaderboard
👩‍💻 Adding a dataset How to add a new task/dataset to MTEB
👩‍💻 Adding a benchmark How to add a new benchmark to MTEB and to the leaderboard
🤝 Contributing How to contribute to MTEB and set it up for development

Popular repositories Loading

  1. mteb mteb Public

    MTEB: State-of-the-art evaluation of embeddings across languages and modalities

    Python 3.4k 692

  2. results results Public

    Data for the MTEB leaderboard

    Python 60 191

  3. leaderboard leaderboard Public archive

    Code for the MTEB leaderboard

    Python 32 15

  4. embedders-dilemma embedders-dilemma Public

    Code, datasets, and results for "The Embedder's Dilemma: LLMs Are Better, but at What Cost?" (COLM 2026). MTEB(LLM) benchmarks 10 LLMs and 26 embedding models on 37 tasks with full cost and through…

    Python 30 1

  5. arena arena Public

    Code for the MTEB Arena

    Python 25 12

  6. mtebpaper mtebpaper Public

    Resources & scripts for the paper "MTEB: Massive Text Embedding Benchmark"

    Python 18 5

Repositories

Showing 10 of 11 repositories
  • mteb Public

    MTEB: State-of-the-art evaluation of embeddings across languages and modalities

    embeddings-benchmark/mteb's past year of commit activity
    Python 3,418 Apache-2.0 692 276 (3 issues need help) 54 Updated Sep 9, 2026
  • MTEB-gym-v2 Public

    Offline LLM-judged arena for embedding models

    embeddings-benchmark/MTEB-gym-v2's past year of commit activity
    Python 2 3 18 1 Updated Sep 8, 2026
  • results Public

    Data for the MTEB leaderboard

    embeddings-benchmark/results's past year of commit activity
    Python 60 CC0-1.0 190 0 6 Updated Sep 6, 2026
  • embeddings-benchmark/leaderboard-frontend's past year of commit activity
    Svelte 2 Apache-2.0 5 1 3 Updated Aug 20, 2026
  • embedders-dilemma Public

    Code, datasets, and results for "The Embedder's Dilemma: LLMs Are Better, but at What Cost?" (COLM 2026). MTEB(LLM) benchmarks 10 LLMs and 26 embedding models on 37 tasks with full cost and throughput accounting.

    embeddings-benchmark/embedders-dilemma's past year of commit activity
    Python 30 Apache-2.0 1 0 0 Updated Aug 19, 2026
  • .github Public
    embeddings-benchmark/.github's past year of commit activity
    0 0 0 0 Updated Aug 18, 2026
  • autoembed Public
    embeddings-benchmark/autoembed's past year of commit activity
    Python 0 1 1 0 Updated Aug 1, 2026
  • arena Public

    Code for the MTEB Arena

    embeddings-benchmark/arena's past year of commit activity
    Python 25 12 25 5 Updated Jul 2, 2025
  • miebpaper Public
    embeddings-benchmark/miebpaper's past year of commit activity
    Jupyter Notebook 2 0 0 0 Updated Feb 28, 2025
  • leaderboard Public archive

    Code for the MTEB leaderboard

    embeddings-benchmark/leaderboard's past year of commit activity
    Python 32 15 15 2 Updated Feb 4, 2025