MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Haohang, Lu, Xuan, Su, Mingyi, Zhang, Xuan, Jiang, Ziyan, Nie, Ping, Zou, Kai, Pfister, Tomas, Chen, Wenhu, Zhang, Wei, Shen, Xiaoyu, Meng, Rui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Reasoning in Document Ranking: Why Chain-of-Thought Falls Short
by: Lu, Xuan, et al.
Published: (2025)
by: Lu, Xuan, et al.
Published: (2025)
Tools are under-documented: Simple Document Expansion Boosts Tool Retrieval
by: Lu, Xuan, et al.
Published: (2025)
by: Lu, Xuan, et al.
Published: (2025)
Beyond Global Similarity: Towards Fine-Grained, Multi-Condition Multimodal Retrieval
by: Lu, Xuan, et al.
Published: (2026)
by: Lu, Xuan, et al.
Published: (2026)
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
by: Li, Zhuofeng, et al.
Published: (2026)
by: Li, Zhuofeng, et al.
Published: (2026)
MAGMaR Shared Task System Description: Video Retrieval with OmniEmbed
by: Zhan, Jiaqi Samantha, et al.
Published: (2025)
by: Zhan, Jiaqi Samantha, et al.
Published: (2025)
BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent
by: Chen, Zijian, et al.
Published: (2025)
by: Chen, Zijian, et al.
Published: (2025)
Beyond Content Relevance: Evaluating Instruction Following in Retrieval Models
by: Zhou, Jianqun, et al.
Published: (2024)
by: Zhou, Jianqun, et al.
Published: (2024)
Unifying Multimodal Retrieval via Document Screenshot Embedding
by: Ma, Xueguang, et al.
Published: (2024)
by: Ma, Xueguang, et al.
Published: (2024)
MultiConIR: Towards multi-condition Information Retrieval
by: Lu, Xuan, et al.
Published: (2025)
by: Lu, Xuan, et al.
Published: (2025)
Bridging Modality Gaps in e-Commerce Products via Vision-Language Alignment
by: Zhang, Yipeng, et al.
Published: (2025)
by: Zhang, Yipeng, et al.
Published: (2025)
Dewey Long Context Embedding Model: A Technical Report
by: Zhang, Dun, et al.
Published: (2025)
by: Zhang, Dun, et al.
Published: (2025)
SAIL-Embedding Technical Report: Omni-modal Embedding Foundation Model
by: Lin, Lin, et al.
Published: (2025)
by: Lin, Lin, et al.
Published: (2025)
RREH: Reconstruction Relations Embedded Hashing for Semi-Paired Cross-Modal Retrieval
by: Wang, Jianzong, et al.
Published: (2024)
by: Wang, Jianzong, et al.
Published: (2024)
Are ID Embeddings Necessary? Whitening Pre-trained Text Embeddings for Effective Sequential Recommendation
by: Zhang, Lingzi, et al.
Published: (2024)
by: Zhang, Lingzi, et al.
Published: (2024)
Efficient and High-Fidelity Omni Modality Retrieval
by: Huynh, Chuong, et al.
Published: (2026)
by: Huynh, Chuong, et al.
Published: (2026)
Seed-Guided Topic Discovery with Out-of-Vocabulary Seeds
by: Zhang, Yu, et al.
Published: (2022)
by: Zhang, Yu, et al.
Published: (2022)
An Efficient Framework for Whole-Page Reranking via Single-Modal Supervision
by: Zhang, Zishuai, et al.
Published: (2025)
by: Zhang, Zishuai, et al.
Published: (2025)
MDiffFR: Modality-Guided Diffusion Generation for Cold-start Items in Federated Recommendation
by: Fu, Kang, et al.
Published: (2025)
by: Fu, Kang, et al.
Published: (2025)
Diffusion Cross-domain Recommendation
by: Xuan, Yuner
Published: (2024)
by: Xuan, Yuner
Published: (2024)
SSEmb: A Joint Structural and Semantic Embedding Framework for Mathematical Formula Retrieval
by: Li, Ruyin, et al.
Published: (2025)
by: Li, Ruyin, et al.
Published: (2025)
Towards Cross-Modal Text-Molecule Retrieval with Better Modality Alignment
by: Song, Jia, et al.
Published: (2024)
by: Song, Jia, et al.
Published: (2024)
A Text is Worth Several Tokens: Text Embedding from LLMs Secretly Aligns Well with The Key Tokens
by: Nie, Zhijie, et al.
Published: (2024)
by: Nie, Zhijie, et al.
Published: (2024)
Single-Branch Network Architectures to Close the Modality Gap in Multimodal Recommendation
by: Ganhör, Christian, et al.
Published: (2025)
by: Ganhör, Christian, et al.
Published: (2025)
HySim-LLM: Embedding-Weighted Fine-Tuning Bounds and Manifold Denoising for Domain-Adapted LLMs
by: Jaberi-Douraki, Majid, et al.
Published: (2025)
by: Jaberi-Douraki, Majid, et al.
Published: (2025)
RoarGraph: A Projected Bipartite Graph for Efficient Cross-Modal Approximate Nearest Neighbor Search
by: Chen, Meng, et al.
Published: (2024)
by: Chen, Meng, et al.
Published: (2024)
Closing the Modality Gap for Mixed Modality Search
by: Li, Binxu, et al.
Published: (2025)
by: Li, Binxu, et al.
Published: (2025)
QARM V2: Quantitative Alignment Multi-Modal Recommendation for Reasoning User Sequence Modeling
by: Xia, Tian, et al.
Published: (2026)
by: Xia, Tian, et al.
Published: (2026)
Embedding Compression in Recommender Systems: A Survey
by: Li, Shiwei, et al.
Published: (2024)
by: Li, Shiwei, et al.
Published: (2024)
UniRAG: Universal Retrieval Augmentation for Large Vision Language Models
by: Sharifymoghaddam, Sahel, et al.
Published: (2024)
by: Sharifymoghaddam, Sahel, et al.
Published: (2024)
Leveraging LLMs to Evaluate Usefulness of Document
by: Wang, Xingzhu, et al.
Published: (2025)
by: Wang, Xingzhu, et al.
Published: (2025)
LLMs Meet Isolation Kernel: Lightweight, Learning-free Binary Embeddings for Fast Retrieval
by: Zhang, Zhibo, et al.
Published: (2026)
by: Zhang, Zhibo, et al.
Published: (2026)
LLM-based Semantic Search for Conversational Queries in E-commerce
by: Siddiqui, Emad, et al.
Published: (2026)
by: Siddiqui, Emad, et al.
Published: (2026)
BookGPT: A General Framework for Book Recommendation Empowered by Large Language Model
by: Zhiyuli, Aakas, et al.
Published: (2023)
by: Zhiyuli, Aakas, et al.
Published: (2023)
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval
by: Kong, Fanheng, et al.
Published: (2025)
by: Kong, Fanheng, et al.
Published: (2025)
Enhancing SPARQL Generation by Triplet-order-sensitive Pre-training
by: Su, Chang, et al.
Published: (2024)
by: Su, Chang, et al.
Published: (2024)
S2G-RAG: Structured Sufficiency and Gap Judging for Iterative Retrieval-Augmented QA
by: Li, Minghan, et al.
Published: (2026)
by: Li, Minghan, et al.
Published: (2026)
Improving the Consistency in Cross-Lingual Cross-Modal Retrieval with 1-to-K Contrastive Learning
by: Nie, Zhijie, et al.
Published: (2024)
by: Nie, Zhijie, et al.
Published: (2024)
Bridging the Modality Gap: Dimension Information Alignment and Sparse Spatial Constraint for Image-Text Matching
by: Ma, Xiang, et al.
Published: (2024)
by: Ma, Xiang, et al.
Published: (2024)
Beyond Semantic Similarity: Rethinking Retrieval for Agentic Search via Direct Corpus Interaction
by: Li, Zhuofeng, et al.
Published: (2026)
by: Li, Zhuofeng, et al.
Published: (2026)
Are Multimodal Embeddings Truly Beneficial for Recommendation? A Deep Dive into Whole vs. Individual Modalities
by: Ye, Yu, et al.
Published: (2025)
by: Ye, Yu, et al.
Published: (2025)
Similar Items
-
Rethinking Reasoning in Document Ranking: Why Chain-of-Thought Falls Short
by: Lu, Xuan, et al.
Published: (2025) -
Tools are under-documented: Simple Document Expansion Boosts Tool Retrieval
by: Lu, Xuan, et al.
Published: (2025) -
Beyond Global Similarity: Towards Fine-Grained, Multi-Condition Multimodal Retrieval
by: Lu, Xuan, et al.
Published: (2026) -
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis
by: Li, Zhuofeng, et al.
Published: (2026) -
MAGMaR Shared Task System Description: Video Retrieval with OmniEmbed
by: Zhan, Jiaqi Samantha, et al.
Published: (2025)