Saved in:
| Main Authors: | Iturra-Bocaz, Gabriel, Vo, Danny, Galuscakova, Petra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.23191 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Reproducibility Study of Metacognitive Retrieval-Augmented Generation
by: Iturra-Bocaz, Gabriel, et al.
Published: (2026)
by: Iturra-Bocaz, Gabriel, et al.
Published: (2026)
UiS-IAI@LiveRAG: Retrieval-Augmented Information Nugget-Based Generation of Responses
by: Łajewska, Weronika, et al.
Published: (2025)
by: Łajewska, Weronika, et al.
Published: (2025)
Exploring Large Language Models for Relevance Judgments in Tetun
by: de Jesus, Gabriel, et al.
Published: (2024)
by: de Jesus, Gabriel, et al.
Published: (2024)
LLMJudge: LLMs for Relevance Judgments
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
An Exam-based Evaluation Approach Beyond Traditional Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2024)
by: Farzi, Naghmeh, et al.
Published: (2024)
Benchmarking LLM-based Relevance Judgment Methods
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
Don't Use LLMs to Make Relevance Judgments
by: Soboroff, Ian
Published: (2024)
by: Soboroff, Ian
Published: (2024)
Contextual Relevance and Adaptive Sampling for LLM-Based Document Reranking
by: Huang, Jerry, et al.
Published: (2025)
by: Huang, Jerry, et al.
Published: (2025)
Variations in Relevance Judgments and the Shelf Life of Test Collections
by: Parry, Andrew, et al.
Published: (2025)
by: Parry, Andrew, et al.
Published: (2025)
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Navigating Speech Recording Collections with AI-Generated Illustrations
by: Håland, Sirina, et al.
Published: (2025)
by: Håland, Sirina, et al.
Published: (2025)
A Study of Relevance Judgments.
by: Cuadra, Carlos A.
Published: (1968)
by: Cuadra, Carlos A.
Published: (1968)
LLMs Can Patch Up Missing Relevance Judgments in Evaluation
by: Upadhyay, Shivani, et al.
Published: (2024)
by: Upadhyay, Shivani, et al.
Published: (2024)
LLM4Rerank: LLM-based Auto-Reranking Framework for Recommendations
by: Gao, Jingtong, et al.
Published: (2024)
by: Gao, Jingtong, et al.
Published: (2024)
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Prism-Reranker: Beyond Relevance Scoring -- Jointly Producing Contributions and Evidence for Agentic Retrieval
by: Zhang, Dun
Published: (2026)
by: Zhang, Dun
Published: (2026)
Leveraging Large Language Models for Relevance Judgments in Legal Case Retrieval
by: Ma, Shengjie, et al.
Published: (2024)
by: Ma, Shengjie, et al.
Published: (2024)
Comparative Analysis of Lion and AdamW Optimizers for Cross-Encoder Reranking with MiniLM, GTE, and ModernBERT
by: Kumar, Shahil, et al.
Published: (2025)
by: Kumar, Shahil, et al.
Published: (2025)
TRUE: A Reproducible Framework for LLM-Driven Relevance Judgment in Information Retrieval
by: Dewan, Mouly, et al.
Published: (2025)
by: Dewan, Mouly, et al.
Published: (2025)
ReFIT: Relevance Feedback from a Reranker during Inference
by: Reddy, Revanth Gangi, et al.
Published: (2023)
by: Reddy, Revanth Gangi, et al.
Published: (2023)
Calibration-Disentangled Learning and Relevance-Prioritized Reranking for Calibrated Sequential Recommendation
by: Jeon, Hyunsik, et al.
Published: (2024)
by: Jeon, Hyunsik, et al.
Published: (2024)
Next-Scale Generative Reranking: A Tree-based Generative Rerank Method at Meituan
by: Wang, Shuli, et al.
Published: (2026)
by: Wang, Shuli, et al.
Published: (2026)
Can We Use Large Language Models to Fill Relevance Judgment Holes?
by: Abbasiantaeb, Zahra, et al.
Published: (2024)
by: Abbasiantaeb, Zahra, et al.
Published: (2024)
Rerank Before You Reason: Analyzing Reranking Tradeoffs through Effective Token Cost in Deep Search Agents
by: Sharifymoghaddam, Sahel, et al.
Published: (2026)
by: Sharifymoghaddam, Sahel, et al.
Published: (2026)
Formalized Information Needs Improve Large-Language-Model Relevance Judgments
by: Keller, Jüri, et al.
Published: (2026)
by: Keller, Jüri, et al.
Published: (2026)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
by: Abdallah, Abdelrahman, et al.
Published: (2025)
by: Abdallah, Abdelrahman, et al.
Published: (2025)
Query-driven Relevant Paragraph Extraction from Legal Judgments
by: Santosh, T. Y. S. S, et al.
Published: (2024)
by: Santosh, T. Y. S. S, et al.
Published: (2024)
How Relevance Emerges: Interpreting LoRA Fine-Tuning in Reranking LLMs
by: Nijasure, Atharva, et al.
Published: (2025)
by: Nijasure, Atharva, et al.
Published: (2025)
ToolRerank: Adaptive and Hierarchy-Aware Reranking for Tool Retrieval
by: Zheng, Yuanhang, et al.
Published: (2024)
by: Zheng, Yuanhang, et al.
Published: (2024)
Aligning Query Representation with Rewritten Query and Relevance Judgments in Conversational Search
by: Mo, Fengran, et al.
Published: (2024)
by: Mo, Fengran, et al.
Published: (2024)
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
The Effect of Document Summarization on LLM-Based Relevance Judgments
by: Mohtadi, Samaneh, et al.
Published: (2025)
by: Mohtadi, Samaneh, et al.
Published: (2025)
Multimodal Item Scoring for Natural Language Recommendation via Gaussian Process Regression with LLM Relevance Judgments
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
EviRerank: Adaptive Evidence Construction for Long-Document LLM Reranking
by: Li, Minghan, et al.
Published: (2024)
by: Li, Minghan, et al.
Published: (2024)
Mitigating the Threshold Priming Effect in Large Language Model-Based Relevance Judgments via Personality Infusing
by: Chen, Nuo, et al.
Published: (2025)
by: Chen, Nuo, et al.
Published: (2025)
LongEval at CLEF 2025: Longitudinal Evaluation of IR Model Performance
by: Cancellieri, Matteo, et al.
Published: (2025)
by: Cancellieri, Matteo, et al.
Published: (2025)
LLM-based Listwise Reranking under the Effect of Positional Bias
by: Qiao, Jingfen, et al.
Published: (2026)
by: Qiao, Jingfen, et al.
Published: (2026)
FGR-ColBERT: Identifying Fine-Grained Relevance Tokens During Retrieval
by: Jarolím, Antonín, et al.
Published: (2026)
by: Jarolím, Antonín, et al.
Published: (2026)
JurisTCU: A Brazilian Portuguese Information Retrieval Dataset with Query Relevance Judgments
by: Fernandes, Leandro Carísio, et al.
Published: (2025)
by: Fernandes, Leandro Carísio, et al.
Published: (2025)
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance Judgment
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
Similar Items
-
A Reproducibility Study of Metacognitive Retrieval-Augmented Generation
by: Iturra-Bocaz, Gabriel, et al.
Published: (2026) -
UiS-IAI@LiveRAG: Retrieval-Augmented Information Nugget-Based Generation of Responses
by: Łajewska, Weronika, et al.
Published: (2025) -
Exploring Large Language Models for Relevance Judgments in Tetun
by: de Jesus, Gabriel, et al.
Published: (2024) -
LLMJudge: LLMs for Relevance Judgments
by: Rahmani, Hossein A., et al.
Published: (2024) -
An Exam-based Evaluation Approach Beyond Traditional Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2024)