MMTEB: Massive Multilingual Text Embedding Benchmark
Fuente:
arXiv
Saved in:
Similar Items
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024)
by: Ciancone, Mathieu, et al.
Published: (2024)
Examining Multilingual Embedding Models Cross-Lingually Through LLM-Generated Adversarial Examples
by: Michail, Andrianos, et al.
Published: (2025)
by: Michail, Andrianos, et al.
Published: (2025)
PARAPHRASUS : A Comprehensive Benchmark for Evaluating Paraphrase Detection Models
by: Michail, Andrianos, et al.
Published: (2024)
by: Michail, Andrianos, et al.
Published: (2024)
Adapting Multilingual Embedding Models to Historical Luxembourgish
by: Michail, Andrianos, et al.
Published: (2025)
by: Michail, Andrianos, et al.
Published: (2025)
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025)
by: Chung, Isaac, et al.
Published: (2025)
Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias
by: Schuhmacher, Elias, et al.
Published: (2026)
by: Schuhmacher, Elias, et al.
Published: (2026)
Interpretable Text Embeddings and Text Similarity Explanation: A Survey
by: Opitz, Juri, et al.
Published: (2025)
by: Opitz, Juri, et al.
Published: (2025)
Sentence Smith: Controllable Edits for Evaluating Text Embeddings
by: Li, Hongji, et al.
Published: (2025)
by: Li, Hongji, et al.
Published: (2025)
BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2026)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2026)
Towards Trustworthy Reranking: A Simple yet Effective Abstention Mechanism
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
Towards a Science of AI Agent Reliability
by: Rabanser, Stephan, et al.
Published: (2026)
by: Rabanser, Stephan, et al.
Published: (2026)
BidirLM: From Text to Omnimodal Bidirectional Encoders by Adapting and Composing Causal LLMs
by: Boizard, Nicolas, et al.
Published: (2026)
by: Boizard, Nicolas, et al.
Published: (2026)
When Does Reasoning Matter? A Controlled Study of Reasoning's Contribution to Model Performance
by: Boizard, Nicolas, et al.
Published: (2025)
by: Boizard, Nicolas, et al.
Published: (2025)
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
CLEF HIPE-2026: Evaluating Accurate and Efficient Person-Place Relation Extraction from Multilingual Historical Texts
by: Opitz, Juri, et al.
Published: (2026)
by: Opitz, Juri, et al.
Published: (2026)
Relações entre satisfação, competência, saúde e absenteísmo no trabalho em uma grande instituição bancária pública
by: Vitor Hugo Bernstorff
Published: (2008)
by: Vitor Hugo Bernstorff
Published: (2008)
MASAI: Modular Architecture for Software-engineering AI Agents
by: Arora, Daman, et al.
Published: (2024)
by: Arora, Daman, et al.
Published: (2024)
SigCLR: Sigmoid Contrastive Learning of Visual Representations
by: Çağatan, Ömer Veysel
Published: (2024)
by: Çağatan, Ömer Veysel
Published: (2024)
UNSEE: Unsupervised Non-contrastive Sentence Embeddings
by: Çağatan, Ömer Veysel
Published: (2024)
by: Çağatan, Ömer Veysel
Published: (2024)
Planarity ranks of modular varieties of semigroups
by: Vladimirovich, Solomatin Denis
Published: (2025)
by: Vladimirovich, Solomatin Denis
Published: (2025)
Automatic Generation and Evaluation of Reading Comprehension Test Items with Large Language Models
by: Säuberli, Andreas, et al.
Published: (2024)
by: Säuberli, Andreas, et al.
Published: (2024)
Should We Still Pretrain Encoders with Masked Language Modeling?
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
A Review of the Challenges with Massive Web-mined Corpora Used in Large Language Models Pre-Training
by: Perełkiewicz, Michał, et al.
Published: (2024)
by: Perełkiewicz, Michał, et al.
Published: (2024)
2D Model for Ca2+$Ca^{2+}$ Dynamics Regulating IP3$IP_3$, ATP and Insulin in A Pancreatic β$\beta$‐Cell
by: Vaishali Vaishali, et al.
Published: (2024)
by: Vaishali Vaishali, et al.
Published: (2024)
Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB
by: Ren, Xingyu, et al.
Published: (2025)
by: Ren, Xingyu, et al.
Published: (2025)
Uncovering RL Integration in SSL Loss: Objective-Specific Implications for Data-Efficient RL
by: Çağatan, Ömer Veysel, et al.
Published: (2024)
by: Çağatan, Ömer Veysel, et al.
Published: (2024)
Failure Modes of Maximum Entropy RLHF
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
by: Çağatan, Ömer Veysel, et al.
Published: (2025)
MIEB: Massive Image Embedding Benchmark
by: Xiao, Chenghao, et al.
Published: (2025)
by: Xiao, Chenghao, et al.
Published: (2025)
Lossless propagation of PT graphene plasmons
by: Sygrimis, Andrianos, et al.
Published: (2026)
by: Sygrimis, Andrianos, et al.
Published: (2026)
ChunkNorris: A High-Performance and Low-Energy Approach to PDF Parsing and Chunking
by: Ciancone, Mathieu, et al.
Published: (2025)
by: Ciancone, Mathieu, et al.
Published: (2025)
Oversight Structures for Agentic AI in Public-Sector Organizations
by: Schmitz, Chris, et al.
Published: (2025)
by: Schmitz, Chris, et al.
Published: (2025)
Exposing Assumptions in AI Benchmarks through Cognitive Modelling
by: Rystrøm, Jonathan H., et al.
Published: (2024)
by: Rystrøm, Jonathan H., et al.
Published: (2024)
OxEnsemble: Fair Ensembles for Low-Data Classification
by: Rystrøm, Jonathan, et al.
Published: (2025)
by: Rystrøm, Jonathan, et al.
Published: (2025)
First case of MPox in Pakistan: What can we learn from it?
by: Areeba Fareed, et al.
Published: (2024)
by: Areeba Fareed, et al.
Published: (2024)
sui-1: Grounded and Verifiable Long-Form Summarization
by: Droste, Benedikt, et al.
Published: (2026)
by: Droste, Benedikt, et al.
Published: (2026)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
by: Idahl, Maximilian, et al.
Published: (2026)
by: Idahl, Maximilian, et al.
Published: (2026)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
by: Zhan, Qiusi, et al.
Published: (2025)
by: Zhan, Qiusi, et al.
Published: (2025)
SMCLM: Semantically Meaningful Causal Language Modeling for Autoregressive Paraphrase Generation
by: Perełkiewicz, Michał, et al.
Published: (2025)
by: Perełkiewicz, Michał, et al.
Published: (2025)
PL-MTEB: Polish Massive Text Embedding Benchmark
by: Poświata, Rafał, et al.
Published: (2024)
by: Poświata, Rafał, et al.
Published: (2024)
HiStruct+: Improving Extractive Text Summarization with Hierarchical Structure Information
by: Ruan, Qian, et al.
Published: (2022)
by: Ruan, Qian, et al.
Published: (2022)
Similar Items
-
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024) -
Examining Multilingual Embedding Models Cross-Lingually Through LLM-Generated Adversarial Examples
by: Michail, Andrianos, et al.
Published: (2025) -
PARAPHRASUS : A Comprehensive Benchmark for Evaluating Paraphrase Detection Models
by: Michail, Andrianos, et al.
Published: (2024) -
Adapting Multilingual Embedding Models to Historical Luxembourgish
by: Michail, Andrianos, et al.
Published: (2025) -
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025)