Overview of the TREC 2022 deep learning track
Fuente:
arXiv
Guardado en:
| Autores principales: | Craswell, Nick, Mitra, Bhaskar, Yilmaz, Emine, Campos, Daniel, Lin, Jimmy, Voorhees, Ellen M., Soboroff, Ian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Overview of the TREC 2023 deep learning track
por: Craswell, Nick, et al.
Publicado: (2025)
por: Craswell, Nick, et al.
Publicado: (2025)
Overview of the TREC 2021 deep learning track
por: Craswell, Nick, et al.
Publicado: (2025)
por: Craswell, Nick, et al.
Publicado: (2025)
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
por: Thakur, Nandan, et al.
Publicado: (2025)
por: Thakur, Nandan, et al.
Publicado: (2025)
Synthetic Test Collections for Retrieval Evaluation
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
por: Pradeep, Ronak, et al.
Publicado: (2024)
por: Pradeep, Ronak, et al.
Publicado: (2024)
Towards Understanding Bias in Synthetic Data for Evaluation
por: Rahmani, Hossein A., et al.
Publicado: (2025)
por: Rahmani, Hossein A., et al.
Publicado: (2025)
Overview of the TREC 2025 Retrieval Augmented Generation (RAG) Track
por: Upadhyay, Shivani, et al.
Publicado: (2026)
por: Upadhyay, Shivani, et al.
Publicado: (2026)
Large language models can accurately predict searcher preferences
por: Thomas, Paul, et al.
Publicado: (2023)
por: Thomas, Paul, et al.
Publicado: (2023)
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework
por: Pradeep, Ronak, et al.
Publicado: (2024)
por: Pradeep, Ronak, et al.
Publicado: (2024)
JudgeBlender: Ensembling Judgments for Automatic Relevance Assessment
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look
por: Upadhyay, Shivani, et al.
Publicado: (2024)
por: Upadhyay, Shivani, et al.
Publicado: (2024)
TREC iKAT 2023: The Interactive Knowledge Assistance Track Overview
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
Overview of the TREC 2025 Tip-of-the-Tongue track
por: Arguello, Jaime, et al.
Publicado: (2026)
por: Arguello, Jaime, et al.
Publicado: (2026)
Lessons from the TREC Plain Language Adaptation of Biomedical Abstracts (PLABA) track
por: Ondov, Brian, et al.
Publicado: (2025)
por: Ondov, Brian, et al.
Publicado: (2025)
SynDL: A Large-Scale Synthetic Test Collection for Passage Retrieval
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
Towards Group-aware Search Success
por: Wu, Haolun, et al.
Publicado: (2024)
por: Wu, Haolun, et al.
Publicado: (2024)
The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
por: Pradeep, Ronak, et al.
Publicado: (2025)
por: Pradeep, Ronak, et al.
Publicado: (2025)
Understanding the Role of User Profile in the Personalization of Large Language Models
por: Wu, Bin, et al.
Publicado: (2024)
por: Wu, Bin, et al.
Publicado: (2024)
TREC iKAT 2023: A Test Collection for Evaluating Conversational and Interactive Knowledge Assistants
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
Overview of the TREC 2025 RAGTIME Track
por: Lawrie, Dawn, et al.
Publicado: (2026)
por: Lawrie, Dawn, et al.
Publicado: (2026)
LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help?
por: Takehi, Rikiya, et al.
Publicado: (2024)
por: Takehi, Rikiya, et al.
Publicado: (2024)
Overview of BioASQ 2022: The tenth BioASQ challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
por: Nentidis, Anastasios, et al.
Publicado: (2022)
por: Nentidis, Anastasios, et al.
Publicado: (2022)
Rethinking Agentic Search with Pi-Serini: Is Lexical Retrieval Sufficient?
por: Hsu, Tz-Huan, et al.
Publicado: (2026)
por: Hsu, Tz-Huan, et al.
Publicado: (2026)
ORBIT: Scalable and Verifiable Data Generation for Search Agents on a Tight Budget
por: Thakur, Nandan, et al.
Publicado: (2026)
por: Thakur, Nandan, et al.
Publicado: (2026)
Hard Negatives, Hard Lessons: Revisiting Training Data Quality for Robust Information Retrieval with LLMs
por: Thakur, Nandan, et al.
Publicado: (2025)
por: Thakur, Nandan, et al.
Publicado: (2025)
Still Fresh? Evaluating Temporal Drift in Retrieval Benchmarks
por: Kuissi, Nathan, et al.
Publicado: (2026)
por: Kuissi, Nathan, et al.
Publicado: (2026)
Sociotechnical Implications of Generative Artificial Intelligence for Information Access
por: Mitra, Bhaskar, et al.
Publicado: (2024)
por: Mitra, Bhaskar, et al.
Publicado: (2024)
Don't Use LLMs to Make Relevance Judgments
por: Soboroff, Ian
Publicado: (2024)
por: Soboroff, Ian
Publicado: (2024)
Leveraging LLMs for Synthesizing Training Data Across Many Languages in Multilingual Dense Retrieval
por: Thakur, Nandan, et al.
Publicado: (2023)
por: Thakur, Nandan, et al.
Publicado: (2023)
A Survey on Retrieval-Augmented Text Generation for Large Language Models
por: Huang, Yizheng, et al.
Publicado: (2024)
por: Huang, Yizheng, et al.
Publicado: (2024)
CoverageBench: Evaluating Information Coverage across Tasks and Domains
por: Samuel, Saron, et al.
Publicado: (2026)
por: Samuel, Saron, et al.
Publicado: (2026)
Overview of the TalentCLEF 2025: Skill and Job Title Intelligence for Human Capital Management
por: Gasco, Luis, et al.
Publicado: (2025)
por: Gasco, Luis, et al.
Publicado: (2025)
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
por: Thakur, Nandan, et al.
Publicado: (2025)
por: Thakur, Nandan, et al.
Publicado: (2025)
LLMJudge: LLMs for Relevance Judgments
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
Judging the Judges: A Collection of LLM-Generated Relevance Judgements
por: Rahmani, Hossein A., et al.
Publicado: (2025)
por: Rahmani, Hossein A., et al.
Publicado: (2025)
Report on the 1st Workshop on Large Language Model for Evaluation in Information Retrieval (LLM4Eval 2024) at SIGIR 2024
por: Rahmani, Hossein A., et al.
Publicado: (2024)
por: Rahmani, Hossein A., et al.
Publicado: (2024)
Image-Seeking Intent Prediction for Cross-Device Product Search
por: Hendriksen, Mariya, et al.
Publicado: (2025)
por: Hendriksen, Mariya, et al.
Publicado: (2025)
Explainability of Text Processing and Retrieval Methods: A Survey
por: Saha, Sourav, et al.
Publicado: (2022)
por: Saha, Sourav, et al.
Publicado: (2022)
NanoKnow: How to Know What Your Language Model Knows
por: Gu, Lingwei, et al.
Publicado: (2026)
por: Gu, Lingwei, et al.
Publicado: (2026)
Learning to Ask: Conversational Product Search via Representation Learning
por: Zou, Jie, et al.
Publicado: (2024)
por: Zou, Jie, et al.
Publicado: (2024)
Ejemplares similares
-
Overview of the TREC 2023 deep learning track
por: Craswell, Nick, et al.
Publicado: (2025) -
Overview of the TREC 2021 deep learning track
por: Craswell, Nick, et al.
Publicado: (2025) -
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
por: Thakur, Nandan, et al.
Publicado: (2025) -
Synthetic Test Collections for Retrieval Evaluation
por: Rahmani, Hossein A., et al.
Publicado: (2024) -
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
por: Pradeep, Ronak, et al.
Publicado: (2024)