MIEB: Massive Image Embedding Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Chenghao, Chung, Isaac, Kerboua, Imene, Stirling, Jamie, Zhang, Xin, Kardos, Márton, Solomatin, Roman, Moubayed, Noura Al, Enevoldsen, Kenneth, Muennighoff, Niklas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025)
by: Chung, Isaac, et al.
Published: (2025)
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
by: Enevoldsen, Kenneth, et al.
Published: (2024)
by: Enevoldsen, Kenneth, et al.
Published: (2024)
HUME: Measuring the Human-Model Performance Gap in Text Embedding Tasks
by: Assadi, Adnan El, et al.
Published: (2025)
by: Assadi, Adnan El, et al.
Published: (2025)
MAEB: Massive Audio Embedding Benchmark
by: Assadi, Adnan El, et al.
Published: (2026)
by: Assadi, Adnan El, et al.
Published: (2026)
RAR-b: Reasoning as Retrieval Benchmark
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
topicwizard -- a Modern, Model-agnostic Framework for Topic Model Visualization and Interpretation
by: Kardos, Márton, et al.
Published: (2025)
by: Kardos, Márton, et al.
Published: (2025)
Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
by: Stirling, Jamie S. J., et al.
Published: (2026)
by: Stirling, Jamie S. J., et al.
Published: (2026)
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2024)
by: Stirling, Jamie, et al.
Published: (2024)
MuLD: The Multitask Long Document Benchmark
by: Hudson, G Thomas, et al.
Published: (2022)
by: Hudson, G Thomas, et al.
Published: (2022)
Naturalistic measure of social norms alignment
by: Kostiuk, Yevhen, et al.
Published: (2026)
by: Kostiuk, Yevhen, et al.
Published: (2026)
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024)
by: Ciancone, Mathieu, et al.
Published: (2024)
Everything is a Video: Unifying Modalities through Next-Frame Prediction
by: Hudson, G. Thomas, et al.
Published: (2024)
by: Hudson, G. Thomas, et al.
Published: (2024)
Early Detection and Reduction of Memorisation for Domain Adaptation and Instruction Tuning
by: Slack, Dean L., et al.
Published: (2025)
by: Slack, Dean L., et al.
Published: (2025)
MMTEB: Massive Multilingual Text Embedding Benchmark
by: Enevoldsen, Kenneth, et al.
Published: (2025)
by: Enevoldsen, Kenneth, et al.
Published: (2025)
Adversarial Defence without Adversarial Defence: Enhancing Language Model Robustness via Instance-level Principal Component Removal
by: Wang, Yang, et al.
Published: (2025)
by: Wang, Yang, et al.
Published: (2025)
$S^3$ -- Semantic Signal Separation
by: Kardos, Márton, et al.
Published: (2024)
by: Kardos, Márton, et al.
Published: (2024)
Topeax -- An Improved Clustering Topic Model with Density Peak Detection and Lexical-Semantic Term Importance
by: Kardos, Márton
Published: (2026)
by: Kardos, Márton
Published: (2026)
AttenCraft: Attention-guided Disentanglement of Multiple Concepts for Text-to-Image Customization
by: Shentu, Junjie, et al.
Published: (2024)
by: Shentu, Junjie, et al.
Published: (2024)
Textual Localization: Decomposing Multi-concept Images for Subject-Driven Text-to-Image Generation
by: Shentu, Junjie, et al.
Published: (2024)
by: Shentu, Junjie, et al.
Published: (2024)
Audio Contrastive-based Fine-tuning: Decoupling Representation Learning and Classification
by: Wang, Yang, et al.
Published: (2023)
by: Wang, Yang, et al.
Published: (2023)
Analyzing LLMs' Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations
by: Xiao, Chenghao, et al.
Published: (2025)
by: Xiao, Chenghao, et al.
Published: (2025)
One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluation
by: Kostiuk, Yevhen, et al.
Published: (2026)
by: Kostiuk, Yevhen, et al.
Published: (2026)
Improving reasoning at inference time via uncertainty minimisation
by: Legrand, Nicolas, et al.
Published: (2026)
by: Legrand, Nicolas, et al.
Published: (2026)
X-ray Made Simple: Lay Radiology Report Generation and Robust Evaluation
by: Zhao, Kun, et al.
Published: (2024)
by: Zhao, Kun, et al.
Published: (2024)
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2026)
by: Stirling, Jamie, et al.
Published: (2026)
Exposing Assumptions in AI Benchmarks through Cognitive Modelling
by: Rystrøm, Jonathan H., et al.
Published: (2024)
by: Rystrøm, Jonathan H., et al.
Published: (2024)
The Power of Next-Frame Prediction for Learning Physical Laws
by: Winterbottom, Thomas, et al.
Published: (2024)
by: Winterbottom, Thomas, et al.
Published: (2024)
Pixel Sentence Representation Learning
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
by: Wu, Siwei, et al.
Published: (2024)
by: Wu, Siwei, et al.
Published: (2024)
Video Prediction of Dynamic Physical Simulations With Pixel-Space Spatiotemporal Transformers
by: Slack, Dean L, et al.
Published: (2025)
by: Slack, Dean L, et al.
Published: (2025)
Disentangling Racial Phenotypes: Fine-Grained Control of Race-related Facial Phenotype Characteristics
by: Yucer, Seyma, et al.
Published: (2024)
by: Yucer, Seyma, et al.
Published: (2024)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
by: Zambrano, Alejandra, et al.
Published: (2026)
by: Zambrano, Alejandra, et al.
Published: (2026)
LLMs for LLMs: A Structured Prompting Methodology for Long Legal Documents
by: Klem, Strahinja, et al.
Published: (2025)
by: Klem, Strahinja, et al.
Published: (2025)
AutoIntent: AutoML for Text Classification
by: Alekseev, Ilya, et al.
Published: (2025)
by: Alekseev, Ilya, et al.
Published: (2025)
C-Pack: Packed Resources For General Chinese Embeddings
by: Xiao, Shitao, et al.
Published: (2023)
by: Xiao, Shitao, et al.
Published: (2023)
Grounding Text Embeddings in Stakeholder Associations
by: Rystrøm, Jonathan, et al.
Published: (2026)
by: Rystrøm, Jonathan, et al.
Published: (2026)
DANSK and DaCy 2.6.0: Domain Generalization of Danish Named Entity Recognition
by: Enevoldsen, Kenneth, et al.
Published: (2024)
by: Enevoldsen, Kenneth, et al.
Published: (2024)
KMMLU: Measuring Massive Multitask Language Understanding in Korean
by: Son, Guijin, et al.
Published: (2024)
by: Son, Guijin, et al.
Published: (2024)
Massive Sound Embedding Benchmark (MSEB)
by: Heigold, Georg, et al.
Published: (2026)
by: Heigold, Georg, et al.
Published: (2026)
Inference-Time Decomposition of Activations (ITDA): A Scalable Approach to Interpreting Large Language Models
by: Leask, Patrick, et al.
Published: (2025)
by: Leask, Patrick, et al.
Published: (2025)
Similar Items
-
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025) -
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
by: Enevoldsen, Kenneth, et al.
Published: (2024) -
HUME: Measuring the Human-Model Performance Gap in Text Embedding Tasks
by: Assadi, Adnan El, et al.
Published: (2025) -
MAEB: Massive Audio Embedding Benchmark
by: Assadi, Adnan El, et al.
Published: (2026) -
RAR-b: Reasoning as Retrieval Benchmark
by: Xiao, Chenghao, et al.
Published: (2024)