Towards Reliable Testing for Multiple Information Retrieval System Comparisons
Fuente:
arXiv
Saved in:
| Main Authors: | Otero, David, Parapar, Javier, Barreiro, Álvaro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation
by: Otero, David, et al.
Published: (2024)
by: Otero, David, et al.
Published: (2024)
Hybrid Pooling with LLMs via Relevance Context Learning
by: Otero, David, et al.
Published: (2026)
by: Otero, David, et al.
Published: (2026)
LLM-Assisted Pseudo-Relevance Feedback
by: Otero, David, et al.
Published: (2026)
by: Otero, David, et al.
Published: (2026)
Beyond Questions: Leveraging ColBERT for Keyphrase Search
by: Gabín, Jorge, et al.
Published: (2024)
by: Gabín, Jorge, et al.
Published: (2024)
Lost in the Evidence? Reproducing Document Position and Context Size Effects in RAG
by: Gabín, Jorge, et al.
Published: (2026)
by: Gabín, Jorge, et al.
Published: (2026)
Enhancing Automatic Keyphrase Labelling with Text-to-Text Transfer Transformer (T5) Architecture: A Framework for Keyphrase Generation and Filtering
by: Gabín, Jorge, et al.
Published: (2024)
by: Gabín, Jorge, et al.
Published: (2024)
VeriCite: Towards Reliable Citations in Retrieval-Augmented Generation via Rigorous Verification
by: Qian, Haosheng, et al.
Published: (2025)
by: Qian, Haosheng, et al.
Published: (2025)
Human-Computer Interaction as a basis for assessing Geographic Information Retrieval Systems.
by: Manuel Enrique Puebla Martínez
Published: (2018)
by: Manuel Enrique Puebla Martínez
Published: (2018)
Measuring Hypothesis Testing Errors in the Evaluation of Retrieval Systems
by: McKechnie, Jack, et al.
Published: (2025)
by: McKechnie, Jack, et al.
Published: (2025)
GenTREC: The First Test Collection Generated by Large Language Models for Evaluating Information Retrieval Systems
by: Türkmen, Mehmet Deniz, et al.
Published: (2025)
by: Türkmen, Mehmet Deniz, et al.
Published: (2025)
Interactions with Generative Information Retrieval Systems
by: Aliannejadi, Mohammad, et al.
Published: (2024)
by: Aliannejadi, Mohammad, et al.
Published: (2024)
A Retrieval Comparison of Six Published Indexes in the Field of Library and Information Science
by: Keen, E. Michael
Published: (1976)
by: Keen, E. Michael
Published: (1976)
STCALIR: Semi-Synthetic Test Collection for Algerian Legal Information Retrieval
by: Hatem, M'hamed Amine, et al.
Published: (2026)
by: Hatem, M'hamed Amine, et al.
Published: (2026)
MultiConIR: Towards multi-condition Information Retrieval
by: Lu, Xuan, et al.
Published: (2025)
by: Lu, Xuan, et al.
Published: (2025)
How Reliable Are Semantic-ID Tokenizer Comparisons in Generative Recommendation?
by: Zhang, Qian, et al.
Published: (2026)
by: Zhang, Qian, et al.
Published: (2026)
Comparison of Information Retrieval Techniques Applied to IT Support Tickets
by: Pereira, Leonardo Santiago Benitez, et al.
Published: (2025)
by: Pereira, Leonardo Santiago Benitez, et al.
Published: (2025)
Vietnamese Legal Information Retrieval in Question-Answering System
by: Ba, Thiem Nguyen, et al.
Published: (2024)
by: Ba, Thiem Nguyen, et al.
Published: (2024)
CROLoss: Towards a Customizable Loss for Retrieval Models in Recommender Systems
by: Tang, Yongxiang, et al.
Published: (2022)
by: Tang, Yongxiang, et al.
Published: (2022)
Training Dense Retrievers with Multiple Positive Passages
by: Wang, Benben, et al.
Published: (2026)
by: Wang, Benben, et al.
Published: (2026)
Towards Reliable Social A/B Testing: Spillover-Contained Clustering with Robust Post-Experiment Analysis
by: Min, Xu, et al.
Published: (2026)
by: Min, Xu, et al.
Published: (2026)
Computerized Information Storage and Retrieval Systems.
by: Azubuike, Abraham A., et al.
Published: (1988)
by: Azubuike, Abraham A., et al.
Published: (1988)
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
by: Oosterhuis, Harrie, et al.
Published: (2024)
by: Oosterhuis, Harrie, et al.
Published: (2024)
Improving Dense Passage Retrieval with Multiple Positive Passages
by: Chang, Shuai
Published: (2025)
by: Chang, Shuai
Published: (2025)
R^3AG: First Workshop on Refined and Reliable Retrieval Augmented Generation
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Fréchet Distance for Offline Evaluation of Information Retrieval Systems with Sparse Labels
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
Robust Information Retrieval
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
Options-Aware Dense Retrieval for Multiple-Choice query Answering
by: Singh, Manish, et al.
Published: (2025)
by: Singh, Manish, et al.
Published: (2025)
Machine Assistant with Reliable Knowledge: Enhancing Student Learning via RAG-based Retrieval
by: Lian, Yongsheng
Published: (2025)
by: Lian, Yongsheng
Published: (2025)
Automatic Classification in Information Retrieval.
by: van Rijsbergen, C. J.
Published: (1978)
by: van Rijsbergen, C. J.
Published: (1978)
LongRetriever: Towards Ultra-Long Sequence based Candidate Retrieval for Recommendation
by: Ren, Qin, et al.
Published: (2025)
by: Ren, Qin, et al.
Published: (2025)
Towards Self-Evolving Agentic Literature Retrieval
by: Du, Yuwen, et al.
Published: (2026)
by: Du, Yuwen, et al.
Published: (2026)
Semantic Search for Information Retrieval
by: Farivar, Kayla
Published: (2025)
by: Farivar, Kayla
Published: (2025)
Information Retrieval with Entity Linking
by: Shehata, Dahlia
Published: (2024)
by: Shehata, Dahlia
Published: (2024)
Generative Information Retrieval Evaluation
by: Alaofi, Marwah, et al.
Published: (2024)
by: Alaofi, Marwah, et al.
Published: (2024)
REANIMATOR: Reanimate Retrieval Test Collections with Extracted and Synthetic Resources
by: Engelmann, Björn, et al.
Published: (2025)
by: Engelmann, Björn, et al.
Published: (2025)
A Survey of Large Language Model Empowered Agents for Recommendation and Search: Towards Next-Generation Information Retrieval
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Enhancing Spectral Knowledge Interrogation: A Reliable Retrieval-Augmented Generative Framework on Large Language Models
by: Liang, Jiheng, et al.
Published: (2024)
by: Liang, Jiheng, et al.
Published: (2024)
Code-Switching Information Retrieval: Benchmarks, Analysis, and the Limits of Current Retrievers
by: Zeng, Qingcheng, et al.
Published: (2026)
by: Zeng, Qingcheng, et al.
Published: (2026)
UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
On the Reliability of User-Centric Evaluation of Conversational Recommender Systems
by: Müller, Michael, et al.
Published: (2026)
by: Müller, Michael, et al.
Published: (2026)
Similar Items
-
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation
by: Otero, David, et al.
Published: (2024) -
Hybrid Pooling with LLMs via Relevance Context Learning
by: Otero, David, et al.
Published: (2026) -
LLM-Assisted Pseudo-Relevance Feedback
by: Otero, David, et al.
Published: (2026) -
Beyond Questions: Leveraging ColBERT for Keyphrase Search
by: Gabín, Jorge, et al.
Published: (2024) -
Lost in the Evidence? Reproducing Document Position and Context Size Effects in RAG
by: Gabín, Jorge, et al.
Published: (2026)