Revisiting Human-vs-LLM judgments using the TREC Podcast Track
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mansour, Watheq, Culpepper, J. Shane, Mackenzie, Joel, Yates, Andrew |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Overview of the TREC 2025 RAGTIME Track
par: Lawrie, Dawn, et autres
Publié: (2026)
par: Lawrie, Dawn, et autres
Publié: (2026)
Overview of the TREC 2024 NeuCLIR Track
par: Lawrie, Dawn, et autres
Publié: (2025)
par: Lawrie, Dawn, et autres
Publié: (2025)
HLTCOE at TREC 2024 NeuCLIR Track
par: Yang, Eugene, et autres
Publié: (2025)
par: Yang, Eugene, et autres
Publié: (2025)
Overview of the TREC 2023 NeuCLIR Track
par: Lawrie, Dawn, et autres
Publié: (2024)
par: Lawrie, Dawn, et autres
Publié: (2024)
Multi-LLM Token Filtering and Routing for Sequential Recommendation
par: Chen, Wuhan, et autres
Publié: (2026)
par: Chen, Wuhan, et autres
Publié: (2026)
On-Device Large Language Models for Sequential Recommendation
par: Xia, Xin, et autres
Publié: (2026)
par: Xia, Xin, et autres
Publié: (2026)
From Questions to Trust Reports: A LLM-IR Framework for the TREC 2025 DRAGUN Track
par: Alwasiak, Ignacy, et autres
Publié: (2026)
par: Alwasiak, Ignacy, et autres
Publié: (2026)
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
par: Thakur, Nandan, et autres
Publié: (2025)
par: Thakur, Nandan, et autres
Publié: (2025)
Overview of the TREC 2025 Retrieval Augmented Generation (RAG) Track
par: Upadhyay, Shivani, et autres
Publié: (2026)
par: Upadhyay, Shivani, et autres
Publié: (2026)
HLTCOE at TREC 2023 NeuCLIR Track
par: Yang, Eugene, et autres
Publié: (2024)
par: Yang, Eugene, et autres
Publié: (2024)
TMU at TREC Clinical Trials Track 2023
par: Lahiri, Aritra Kumar, et autres
Publié: (2024)
par: Lahiri, Aritra Kumar, et autres
Publié: (2024)
The Effects of Demographic Instructions on LLM Personas
par: de Paula, Angel Felipe Magnossão, et autres
Publié: (2025)
par: de Paula, Angel Felipe Magnossão, et autres
Publié: (2025)
Overview of TREC 2025 Biomedical Generative Retrieval (BioGen) Track
par: Gupta, Deepak, et autres
Publié: (2026)
par: Gupta, Deepak, et autres
Publié: (2026)
Overview of TREC 2024 Biomedical Generative Retrieval (BioGen) Track
par: Gupta, Deepak, et autres
Publié: (2024)
par: Gupta, Deepak, et autres
Publié: (2024)
Table Integration in Data Lakes Unleashed: Pairwise Integrability Judgment, Integrable Set Discovery, and Multi-Tuple Conflict Resolution
par: Ji, Daomin, et autres
Publié: (2024)
par: Ji, Daomin, et autres
Publié: (2024)
Overview of the TREC 2025 Tip-of-the-Tongue track
par: Arguello, Jaime, et autres
Publié: (2026)
par: Arguello, Jaime, et autres
Publié: (2026)
How Much Freedom Does An Effectiveness Metric Really Have?
par: Moffat, Alistair, et autres
Publié: (2023)
par: Moffat, Alistair, et autres
Publié: (2023)
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework
par: Pradeep, Ronak, et autres
Publié: (2024)
par: Pradeep, Ronak, et autres
Publié: (2024)
Team IELAB at TREC Clinical Trial Track 2023: Enhancing Clinical Trial Retrieval with Neural Rankers and Large Language Models
par: Zhuang, Shengyao, et autres
Publié: (2024)
par: Zhuang, Shengyao, et autres
Publié: (2024)
How role-play shapes relevance judgment in zero-shot LLM rankers
par: Wang, Yumeng, et autres
Publié: (2025)
par: Wang, Yumeng, et autres
Publié: (2025)
When LLM Judges Inflate Scores: Exploring Overrating in Relevance Assessment
par: Yu, Chuting, et autres
Publié: (2026)
par: Yu, Chuting, et autres
Publié: (2026)
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking
par: Qiao, Yixuan, et autres
Publié: (2022)
par: Qiao, Yixuan, et autres
Publié: (2022)
TREC iKAT 2023: The Interactive Knowledge Assistance Track Overview
par: Aliannejadi, Mohammad, et autres
Publié: (2024)
par: Aliannejadi, Mohammad, et autres
Publié: (2024)
Generative Relevance Feedback and Convergence of Adaptive Re-Ranking: University of Glasgow Terrier Team at TREC DL 2023
par: Parry, Andrew, et autres
Publié: (2024)
par: Parry, Andrew, et autres
Publié: (2024)
LANCER: LLM Reranking for Nugget Coverage
par: Ju, Jia-Huei, et autres
Publié: (2026)
par: Ju, Jia-Huei, et autres
Publié: (2026)
TREC Initiative with Cheshire II.
par: Larson, Ray R.
Publié: (2001)
par: Larson, Ray R.
Publié: (2001)
RALI@TREC iKAT 2024: Achieving Personalization via Retrieval Fusion in Conversational Search
par: Hui, Yuchen, et autres
Publié: (2024)
par: Hui, Yuchen, et autres
Publié: (2024)
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
par: Pradeep, Ronak, et autres
Publié: (2024)
par: Pradeep, Ronak, et autres
Publié: (2024)
HLTCOE at LiveRAG: GPT-Researcher using ColBERT retrieval
par: Duh, Kevin, et autres
Publié: (2025)
par: Duh, Kevin, et autres
Publié: (2025)
LLM-based Listwise Reranking under the Effect of Positional Bias
par: Qiao, Jingfen, et autres
Publié: (2026)
par: Qiao, Jingfen, et autres
Publié: (2026)
Scientometric Analysis of the German IR Community within TREC & CLEF
par: Kruff, A. K., et autres
Publié: (2025)
par: Kruff, A. K., et autres
Publié: (2025)
GenTREC: The First Test Collection Generated by Large Language Models for Evaluating Information Retrieval Systems
par: Türkmen, Mehmet Deniz, et autres
Publié: (2025)
par: Türkmen, Mehmet Deniz, et autres
Publié: (2025)
Corpus-informed Retrieval Augmented Generation of Clarifying Questions
par: Krasakis, Antonios Minas, et autres
Publié: (2024)
par: Krasakis, Antonios Minas, et autres
Publié: (2024)
Constructing Set-Compositional and Negated Representations for First-Stage Ranking
par: Krasakis, Antonios Minas, et autres
Publié: (2025)
par: Krasakis, Antonios Minas, et autres
Publié: (2025)
Cold-Starting Podcast Ads and Promotions with Multi-Task Learning on Spotify
par: Verma, Shivam, et autres
Publié: (2026)
par: Verma, Shivam, et autres
Publié: (2026)
Evaluating Podcast Recommendations with Profile-Aware LLM-as-a-Judge
par: Fabbri, Francesco, et autres
Publié: (2025)
par: Fabbri, Francesco, et autres
Publié: (2025)
Overview of the TREC 2022 deep learning track
par: Craswell, Nick, et autres
Publié: (2025)
par: Craswell, Nick, et autres
Publié: (2025)
Overview of the TREC 2023 deep learning track
par: Craswell, Nick, et autres
Publié: (2025)
par: Craswell, Nick, et autres
Publié: (2025)
Overview of the TREC 2021 deep learning track
par: Craswell, Nick, et autres
Publié: (2025)
par: Craswell, Nick, et autres
Publié: (2025)
Transforming Podcast Preview Generation: From Expert Models to LLM-Based Systems
par: Zhu, Winstead, et autres
Publié: (2025)
par: Zhu, Winstead, et autres
Publié: (2025)
Documents similaires
-
Overview of the TREC 2025 RAGTIME Track
par: Lawrie, Dawn, et autres
Publié: (2026) -
Overview of the TREC 2024 NeuCLIR Track
par: Lawrie, Dawn, et autres
Publié: (2025) -
HLTCOE at TREC 2024 NeuCLIR Track
par: Yang, Eugene, et autres
Publié: (2025) -
Overview of the TREC 2023 NeuCLIR Track
par: Lawrie, Dawn, et autres
Publié: (2024) -
Multi-LLM Token Filtering and Routing for Sequential Recommendation
par: Chen, Wuhan, et autres
Publié: (2026)