Overview of the TREC 2023 deep learning track
Fuente:
arXiv
Saved in:
| Main Authors: | Craswell, Nick, Mitra, Bhaskar, Yilmaz, Emine, Rahmani, Hossein A., Campos, Daniel, Lin, Jimmy, Voorhees, Ellen M., Soboroff, Ian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Overview of the TREC 2022 deep learning track
by: Craswell, Nick, et al.
Published: (2025)
by: Craswell, Nick, et al.
Published: (2025)
Overview of the TREC 2021 deep learning track
by: Craswell, Nick, et al.
Published: (2025)
by: Craswell, Nick, et al.
Published: (2025)
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Synthetic Test Collections for Retrieval Evaluation
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Towards Understanding Bias in Synthetic Data for Evaluation
by: Rahmani, Hossein A., et al.
Published: (2025)
by: Rahmani, Hossein A., et al.
Published: (2025)
Overview of the TREC 2025 Retrieval Augmented Generation (RAG) Track
by: Upadhyay, Shivani, et al.
Published: (2026)
by: Upadhyay, Shivani, et al.
Published: (2026)
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework
by: Pradeep, Ronak, et al.
Published: (2024)
by: Pradeep, Ronak, et al.
Published: (2024)
SynDL: A Large-Scale Synthetic Test Collection for Passage Retrieval
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
by: Thakur, Nandan, et al.
Published: (2025)
by: Thakur, Nandan, et al.
Published: (2025)
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look
by: Upadhyay, Shivani, et al.
Published: (2024)
by: Upadhyay, Shivani, et al.
Published: (2024)
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
by: Pradeep, Ronak, et al.
Published: (2024)
by: Pradeep, Ronak, et al.
Published: (2024)
Overview of the TREC 2025 Tip-of-the-Tongue track
by: Arguello, Jaime, et al.
Published: (2026)
by: Arguello, Jaime, et al.
Published: (2026)
Towards Group-aware Search Success
by: Wu, Haolun, et al.
Published: (2024)
by: Wu, Haolun, et al.
Published: (2024)
LLMJudge: LLMs for Relevance Judgments
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Judging the Judges: A Collection of LLM-Generated Relevance Judgements
by: Rahmani, Hossein A., et al.
Published: (2025)
by: Rahmani, Hossein A., et al.
Published: (2025)
Report on the 1st Workshop on Large Language Model for Evaluation in Information Retrieval (LLM4Eval 2024) at SIGIR 2024
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
by: Pradeep, Ronak, et al.
Published: (2025)
by: Pradeep, Ronak, et al.
Published: (2025)
Large language models can accurately predict searcher preferences
by: Thomas, Paul, et al.
Published: (2023)
by: Thomas, Paul, et al.
Published: (2023)
Overview of the TREC 2025 RAGTIME Track
by: Lawrie, Dawn, et al.
Published: (2026)
by: Lawrie, Dawn, et al.
Published: (2026)
TREC iKAT 2023: The Interactive Knowledge Assistance Track Overview
by: Aliannejadi, Mohammad, et al.
Published: (2024)
by: Aliannejadi, Mohammad, et al.
Published: (2024)
LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help?
by: Takehi, Rikiya, et al.
Published: (2024)
by: Takehi, Rikiya, et al.
Published: (2024)
Understanding the Role of User Profile in the Personalization of Large Language Models
by: Wu, Bin, et al.
Published: (2024)
by: Wu, Bin, et al.
Published: (2024)
Don't Use LLMs to Make Relevance Judgments
by: Soboroff, Ian
Published: (2024)
by: Soboroff, Ian
Published: (2024)
HLTCOE at TREC 2023 NeuCLIR Track
by: Yang, Eugene, et al.
Published: (2024)
by: Yang, Eugene, et al.
Published: (2024)
TMU at TREC Clinical Trials Track 2023
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
Overview of the TREC 2023 NeuCLIR Track
by: Lawrie, Dawn, et al.
Published: (2024)
by: Lawrie, Dawn, et al.
Published: (2024)
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor
by: Upadhyay, Shivani, et al.
Published: (2024)
by: Upadhyay, Shivani, et al.
Published: (2024)
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Overview of the TREC 2024 NeuCLIR Track
by: Lawrie, Dawn, et al.
Published: (2025)
by: Lawrie, Dawn, et al.
Published: (2025)
Adaptive Retrieval-Augmented Generation for Conversational Systems
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
Lessons from the TREC Plain Language Adaptation of Biomedical Abstracts (PLABA) track
by: Ondov, Brian, et al.
Published: (2025)
by: Ondov, Brian, et al.
Published: (2025)
TREC iKAT 2023: A Test Collection for Evaluating Conversational and Interactive Knowledge Assistants
by: Aliannejadi, Mohammad, et al.
Published: (2024)
by: Aliannejadi, Mohammad, et al.
Published: (2024)
Overview of TREC 2024 Biomedical Generative Retrieval (BioGen) Track
by: Gupta, Deepak, et al.
Published: (2024)
by: Gupta, Deepak, et al.
Published: (2024)
Overview of TREC 2025 Biomedical Generative Retrieval (BioGen) Track
by: Gupta, Deepak, et al.
Published: (2026)
by: Gupta, Deepak, et al.
Published: (2026)
Study on LLMs for Promptagator-Style Dense Retriever Training
by: Gwon, Daniel, et al.
Published: (2025)
by: Gwon, Daniel, et al.
Published: (2025)
Search and Society: Reimagining Information Access for Radical Futures
by: Mitra, Bhaskar
Published: (2024)
by: Mitra, Bhaskar
Published: (2024)
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking
by: Qiao, Yixuan, et al.
Published: (2022)
by: Qiao, Yixuan, et al.
Published: (2022)
QueryBuilder: Human-in-the-Loop Query Development for Information Retrieval
by: Kandula, Hemanth, et al.
Published: (2024)
by: Kandula, Hemanth, et al.
Published: (2024)
A Systematic Study of Pseudo-Relevance Feedback with LLMs
by: Jedidi, Nour, et al.
Published: (2026)
by: Jedidi, Nour, et al.
Published: (2026)
Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation
by: He, Xuhong, et al.
Published: (2026)
by: He, Xuhong, et al.
Published: (2026)
Similar Items
-
Overview of the TREC 2022 deep learning track
by: Craswell, Nick, et al.
Published: (2025) -
Overview of the TREC 2021 deep learning track
by: Craswell, Nick, et al.
Published: (2025) -
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment
by: Rahmani, Hossein A., et al.
Published: (2024) -
Synthetic Test Collections for Retrieval Evaluation
by: Rahmani, Hossein A., et al.
Published: (2024) -
Towards Understanding Bias in Synthetic Data for Evaluation
by: Rahmani, Hossein A., et al.
Published: (2025)