Retrieval or Representation? Reassessing Benchmark Gaps in Multilingual and Visually Rich RAG
Fuente:
arXiv
Saved in:
| Main Authors: | Asenov, Martin, Benkirane, Kenza, Goldwater, Dan, Ghodsi, Aneiss |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DISCO: Document Intelligence Suite for COmparative Evaluation
by: Benkirane, Kenza, et al.
Published: (2026)
by: Benkirane, Kenza, et al.
Published: (2026)
LAD-RAG: Layout-aware Dynamic RAG for Visually-Rich Document Understanding
by: Sourati, Zhivar, et al.
Published: (2025)
by: Sourati, Zhivar, et al.
Published: (2025)
Roles of MLLMs in Visually Rich Document Retrieval for RAG: A Survey
by: Zhang, Xiantao
Published: (2025)
by: Zhang, Xiantao
Published: (2025)
VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents
by: Tanaka, Ryota, et al.
Published: (2025)
by: Tanaka, Ryota, et al.
Published: (2025)
RichRAG: Crafting Rich Responses for Multi-faceted Queries in Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries
by: Wu, Yin, et al.
Published: (2025)
by: Wu, Yin, et al.
Published: (2025)
Benchmarking LLM Guardrails in Handling Multilingual Toxicity
by: Yang, Yahan, et al.
Published: (2024)
by: Yang, Yahan, et al.
Published: (2024)
CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG
by: Lee, Nayeon, et al.
Published: (2026)
by: Lee, Nayeon, et al.
Published: (2026)
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
How Retrieved Context Shapes Internal Representations in RAG
by: Yeh, Samuel, et al.
Published: (2026)
by: Yeh, Samuel, et al.
Published: (2026)
Revisiting Common Assumptions about Arabic Dialects in NLP
by: Keleg, Amr, et al.
Published: (2025)
by: Keleg, Amr, et al.
Published: (2025)
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
by: Benkirane, Kenza, et al.
Published: (2024)
by: Benkirane, Kenza, et al.
Published: (2024)
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems
by: Cirillo, Stefano, et al.
Published: (2026)
by: Cirillo, Stefano, et al.
Published: (2026)
YpathRAG:A Retrieval-Augmented Generation Framework and Benchmark for Pathology
by: Yu, Deshui, et al.
Published: (2025)
by: Yu, Deshui, et al.
Published: (2025)
All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG
by: Wang, Dan, et al.
Published: (2026)
by: Wang, Dan, et al.
Published: (2026)
A Grounded Typology of Word Classes
by: Haley, Coleman, et al.
Published: (2024)
by: Haley, Coleman, et al.
Published: (2024)
HyperbolicRAG: Enhancing Retrieval-Augmented Generation with Hyperbolic Representations
by: Cao, Linxiao, et al.
Published: (2025)
by: Cao, Linxiao, et al.
Published: (2025)
MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries
by: Tang, Yixuan, et al.
Published: (2024)
by: Tang, Yixuan, et al.
Published: (2024)
Mind the Gap... or Not? How Translation Errors and Evaluation Details Skew Multilingual Results
by: Peter, Jan-Thorsten, et al.
Published: (2025)
by: Peter, Jan-Thorsten, et al.
Published: (2025)
Investigating Language Preference of Multilingual RAG Systems
by: Park, Jeonghyun, et al.
Published: (2025)
by: Park, Jeonghyun, et al.
Published: (2025)
A framework for analyzing concept representations in neural models
by: Naowarat, Burin, et al.
Published: (2026)
by: Naowarat, Burin, et al.
Published: (2026)
Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets
by: Keleg, Amr, et al.
Published: (2024)
by: Keleg, Amr, et al.
Published: (2024)
Smart-Hiring: An Explainable end-to-end Pipeline for CV Information Extraction and Job Matching
by: Khelkhal, Kenza, et al.
Published: (2025)
by: Khelkhal, Kenza, et al.
Published: (2025)
MultiZebraLogic: A Multilingual Logical Reasoning Benchmark
by: Bruun, Sofie Helene, et al.
Published: (2025)
by: Bruun, Sofie Helene, et al.
Published: (2025)
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models
by: Lyu, Yuanjie, et al.
Published: (2024)
by: Lyu, Yuanjie, et al.
Published: (2024)
Multilingual Previously Fact-Checked Claim Retrieval
by: Pikuliak, Matúš, et al.
Published: (2023)
by: Pikuliak, Matúš, et al.
Published: (2023)
Effective Context in Neural Speech Models
by: Meng, Yen, et al.
Published: (2025)
by: Meng, Yen, et al.
Published: (2025)
Bridging Language Gaps: Advances in Cross-Lingual Information Retrieval with Multilingual LLMs
by: Goworek, Roksana, et al.
Published: (2025)
by: Goworek, Roksana, et al.
Published: (2025)
FinRAGBench-V: A Benchmark for Multimodal RAG with Visual Citation in the Financial Domain
by: Zhao, Suifeng, et al.
Published: (2025)
by: Zhao, Suifeng, et al.
Published: (2025)
Orthogonality and isotropy of speaker and phonetic information in self-supervised speech representations
by: Mohamed, Mukhtar, et al.
Published: (2024)
by: Mohamed, Mukhtar, et al.
Published: (2024)
Multilingual Hallucination Gaps in Large Language Models
by: Chataigner, Cléa, et al.
Published: (2024)
by: Chataigner, Cléa, et al.
Published: (2024)
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG
by: Wang, Wenbin, et al.
Published: (2025)
by: Wang, Wenbin, et al.
Published: (2025)
CRAG -- Comprehensive RAG Benchmark
by: Yang, Xiao, et al.
Published: (2024)
by: Yang, Xiao, et al.
Published: (2024)
Multilingual Retrieval Augmented Generation for Culturally-Sensitive Tasks: A Benchmark for Cross-lingual Robustness
by: Li, Bryan, et al.
Published: (2024)
by: Li, Bryan, et al.
Published: (2024)
Toward Optimal Search and Retrieval for RAG
by: Leto, Alexandria, et al.
Published: (2024)
by: Leto, Alexandria, et al.
Published: (2024)
MultiVENT 2.0: A Massive Multilingual Benchmark for Event-Centric Video Retrieval
by: Kriz, Reno, et al.
Published: (2024)
by: Kriz, Reno, et al.
Published: (2024)
A Systematic Review of Key Retrieval-Augmented Generation (RAG) Systems: Progress, Gaps, and Future Directions
by: Oche, Agada Joseph, et al.
Published: (2025)
by: Oche, Agada Joseph, et al.
Published: (2025)
AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG
by: You, Qijie, et al.
Published: (2026)
by: You, Qijie, et al.
Published: (2026)
TP-RAG: Benchmarking Retrieval-Augmented Large Language Model Agents for Spatiotemporal-Aware Travel Planning
by: Ni, Hang, et al.
Published: (2025)
by: Ni, Hang, et al.
Published: (2025)
Magic Mushroom: A Customizable Benchmark for Fine-grained Analysis of Retrieval Noise Erosion in RAG Systems
by: Zhang, Yuxin, et al.
Published: (2025)
by: Zhang, Yuxin, et al.
Published: (2025)
Similar Items
-
DISCO: Document Intelligence Suite for COmparative Evaluation
by: Benkirane, Kenza, et al.
Published: (2026) -
LAD-RAG: Layout-aware Dynamic RAG for Visually-Rich Document Understanding
by: Sourati, Zhivar, et al.
Published: (2025) -
Roles of MLLMs in Visually Rich Document Retrieval for RAG: A Survey
by: Zhang, Xiantao
Published: (2025) -
VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents
by: Tanaka, Ryota, et al.
Published: (2025) -
RichRAG: Crafting Rich Responses for Multi-faceted Queries in Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)