SARA: A Collection of Sensitivity-Aware Relevance Assessments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | McKechnie, Jack, McDonald, Graham |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Measuring Hypothesis Testing Errors in the Evaluation of Retrieval Systems
von: McKechnie, Jack, et al.
Veröffentlicht: (2025)
von: McKechnie, Jack, et al.
Veröffentlicht: (2025)
Who Benefits from RAG? The Role of Exposure, Utility and Attribution Bias
von: Dehghan, Mahdi, et al.
Veröffentlicht: (2026)
von: Dehghan, Mahdi, et al.
Veröffentlicht: (2026)
Query Exposure Prediction for Groups of Documents in Rankings
von: Jaenich, Thomas, et al.
Veröffentlicht: (2024)
von: Jaenich, Thomas, et al.
Veröffentlicht: (2024)
Document Similarity Enhanced IPS Estimation for Unbiased Learning to Rank
von: Liang, Zeyan, et al.
Veröffentlicht: (2025)
von: Liang, Zeyan, et al.
Veröffentlicht: (2025)
Temporal Fact Conflicts in LLMs: Reproducibility Insights from Unifying DYNAMICQA and MULAN
von: Dey, Ritajit, et al.
Veröffentlicht: (2026)
von: Dey, Ritajit, et al.
Veröffentlicht: (2026)
Quantifying Query Fairness Under Unawareness
von: Jaenich, Thomas, et al.
Veröffentlicht: (2025)
von: Jaenich, Thomas, et al.
Veröffentlicht: (2025)
The Cranfield II Relevance Assessments: A Critical Evaluation
von: Harter, Stephen P.
Veröffentlicht: (1971)
von: Harter, Stephen P.
Veröffentlicht: (1971)
Judging the Judges: A Collection of LLM-Generated Relevance Judgements
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2025)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2025)
Variations in Relevance Judgments and the Shelf Life of Test Collections
von: Parry, Andrew, et al.
Veröffentlicht: (2025)
von: Parry, Andrew, et al.
Veröffentlicht: (2025)
Behind Closed Doors: An Exploratory Study of the Perceptions of Librarians and the Hidden Intellectual Work of Collection Development in Canadian Public Libraries.
von: Nilsen, Kirsti, et al.
Veröffentlicht: (2002)
von: Nilsen, Kirsti, et al.
Veröffentlicht: (2002)
Axiomatic Causal Interventions for Reverse Engineering Relevance Computation in Neural Retrieval Models
von: Chen, Catherine, et al.
Veröffentlicht: (2024)
von: Chen, Catherine, et al.
Veröffentlicht: (2024)
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
Batched Self-Consistency Improves LLM Relevance Assessment and Ranking
von: Korikov, Anton, et al.
Veröffentlicht: (2025)
von: Korikov, Anton, et al.
Veröffentlicht: (2025)
When LLM Judges Inflate Scores: Exploring Overrating in Relevance Assessment
von: Yu, Chuting, et al.
Veröffentlicht: (2026)
von: Yu, Chuting, et al.
Veröffentlicht: (2026)
Searching Personal Collections
von: Bendersky, Michael, et al.
Veröffentlicht: (2024)
von: Bendersky, Michael, et al.
Veröffentlicht: (2024)
Self-Service Circulation: An Exploratory Study.
von: Carey, Robert F., et al.
Veröffentlicht: (1998)
von: Carey, Robert F., et al.
Veröffentlicht: (1998)
SelRoute: Query-Type-Aware Routing for Long-Term Conversational Memory Retrieval
von: McKee, Matthew
Veröffentlicht: (2026)
von: McKee, Matthew
Veröffentlicht: (2026)
Limitations of Automatic Relevance Assessments with Large Language Models for Fair and Reliable Retrieval Evaluation
von: Otero, David, et al.
Veröffentlicht: (2024)
von: Otero, David, et al.
Veröffentlicht: (2024)
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance Judgment
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2025)
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2025)
Calibration-Disentangled Learning and Relevance-Prioritized Reranking for Calibrated Sequential Recommendation
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2024)
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2024)
Beyond Relevance: Evaluate and Improve Retrievers on Perspective Awareness
von: Zhao, Xinran, et al.
Veröffentlicht: (2024)
von: Zhao, Xinran, et al.
Veröffentlicht: (2024)
SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression
von: Jin, Yiqiao, et al.
Veröffentlicht: (2025)
von: Jin, Yiqiao, et al.
Veröffentlicht: (2025)
Judging with Personality and Confidence: A Study on Personality-Conditioned LLM Relevance Assessment
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
von: Chen, Nuo, et al.
Veröffentlicht: (2026)
SPECTRA: Synthetic IR Test Collections with Relevance Oracles and Controlled Distractor Diagnostics
von: Liang, Eric
Veröffentlicht: (2026)
von: Liang, Eric
Veröffentlicht: (2026)
All Eyes on the Ranker: Participatory Auditing to Surface Blind Spots in Ranked Search Results
von: Rezk, Anna Marie, et al.
Veröffentlicht: (2026)
von: Rezk, Anna Marie, et al.
Veröffentlicht: (2026)
LLMJudge: LLMs for Relevance Judgments
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
Generalized Pseudo-Relevance Feedback
von: Tu, Yiteng, et al.
Veröffentlicht: (2025)
von: Tu, Yiteng, et al.
Veröffentlicht: (2025)
LLM-based Relevance Assessment for Web-Scale Search Evaluation at Pinterest
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look
von: Upadhyay, Shivani, et al.
Veröffentlicht: (2024)
von: Upadhyay, Shivani, et al.
Veröffentlicht: (2024)
A Deep Learning Approach for Selective Relevance Feedback
von: Datta, Suchana, et al.
Veröffentlicht: (2024)
von: Datta, Suchana, et al.
Veröffentlicht: (2024)
REAR: A Relevance-Aware Retrieval-Augmented Framework for Open-Domain Question Answering
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
LLM-Assisted Pseudo-Relevance Feedback
von: Otero, David, et al.
Veröffentlicht: (2026)
von: Otero, David, et al.
Veröffentlicht: (2026)
LUCid: Redefining Relevance For Lifelong Personalization
von: Okite, Chimaobi, et al.
Veröffentlicht: (2026)
von: Okite, Chimaobi, et al.
Veröffentlicht: (2026)
R3A: Reinforced Reasoning for Relevance Assessment for RAG in User-Generated Content Platforms
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2025)
von: Yuan, Xiaowei, et al.
Veröffentlicht: (2025)
NeuCLIRBench: A Modern Evaluation Collection for Monolingual, Cross-Language, and Multilingual Information Retrieval
von: Lawrie, Dawn, et al.
Veröffentlicht: (2025)
von: Lawrie, Dawn, et al.
Veröffentlicht: (2025)
Augmenting Sequential Recommendation with Balanced Relevance and Diversity
von: Dang, Yizhou, et al.
Veröffentlicht: (2024)
von: Dang, Yizhou, et al.
Veröffentlicht: (2024)
Generative Retrieval Meets Multi-Graded Relevance
von: Tang, Yubao, et al.
Veröffentlicht: (2024)
von: Tang, Yubao, et al.
Veröffentlicht: (2024)
Don't Use LLMs to Make Relevance Judgments
von: Soboroff, Ian
Veröffentlicht: (2024)
von: Soboroff, Ian
Veröffentlicht: (2024)
Is Relevance Propagated from Retriever to Generator in RAG?
von: Tian, Fangzheng, et al.
Veröffentlicht: (2025)
von: Tian, Fangzheng, et al.
Veröffentlicht: (2025)
Migrating a Job Search Relevance Function
von: Mountain, Bennett, et al.
Veröffentlicht: (2025)
von: Mountain, Bennett, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Measuring Hypothesis Testing Errors in the Evaluation of Retrieval Systems
von: McKechnie, Jack, et al.
Veröffentlicht: (2025) -
Who Benefits from RAG? The Role of Exposure, Utility and Attribution Bias
von: Dehghan, Mahdi, et al.
Veröffentlicht: (2026) -
Query Exposure Prediction for Groups of Documents in Rankings
von: Jaenich, Thomas, et al.
Veröffentlicht: (2024) -
Document Similarity Enhanced IPS Estimation for Unbiased Learning to Rank
von: Liang, Zeyan, et al.
Veröffentlicht: (2025) -
Temporal Fact Conflicts in LLMs: Reproducibility Insights from Unifying DYNAMICQA and MULAN
von: Dey, Ritajit, et al.
Veröffentlicht: (2026)