Evaluation of Temporal Change in IR Test Collections
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Keller, Jüri, Breuer, Timo, Schaer, Philipp |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Replicability Measures for Longitudinal Information Retrieval Evaluation
von: Keller, Jüri, et al.
Veröffentlicht: (2024)
von: Keller, Jüri, et al.
Veröffentlicht: (2024)
Evaluating Contrastive Feedback for Effective User Simulations
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025)
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025)
Formalized Information Needs Improve Large-Language-Model Relevance Judgments
von: Keller, Jüri, et al.
Veröffentlicht: (2026)
von: Keller, Jüri, et al.
Veröffentlicht: (2026)
Context-Driven Interactive Query Simulations Based on Generative Large Language Models
von: Engelmann, Björn, et al.
Veröffentlicht: (2023)
von: Engelmann, Björn, et al.
Veröffentlicht: (2023)
Sim4IA-Bench: A User Simulation Benchmark Suite for Next Query and Utterance Prediction
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025)
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025)
Simplified Longitudinal Retrieval Experiments: A Case Study on Query Expansion and Document Boosting
von: Keller, Jüri, et al.
Veröffentlicht: (2025)
von: Keller, Jüri, et al.
Veröffentlicht: (2025)
LongEval at CLEF 2025: Longitudinal Evaluation of IR Model Performance
von: Cancellieri, Matteo, et al.
Veröffentlicht: (2025)
von: Cancellieri, Matteo, et al.
Veröffentlicht: (2025)
Counterfactual Query Rewriting to Use Historical Relevance Feedback
von: Keller, Jüri, et al.
Veröffentlicht: (2025)
von: Keller, Jüri, et al.
Veröffentlicht: (2025)
LongEval at CLEF 2025: Longitudinal Evaluation of IR Systems on Web and Scientific Data
von: Cancellieri, Matteo, et al.
Veröffentlicht: (2025)
von: Cancellieri, Matteo, et al.
Veröffentlicht: (2025)
Data Fusion of Synthetic Query Variants With Generative Large Language Models
von: Breuer, Timo
Veröffentlicht: (2024)
von: Breuer, Timo
Veröffentlicht: (2024)
REANIMATOR: Reanimate Retrieval Test Collections with Extracted and Synthetic Resources
von: Engelmann, Björn, et al.
Veröffentlicht: (2025)
von: Engelmann, Björn, et al.
Veröffentlicht: (2025)
Second SIGIR Workshop on Simulations for Information Access (Sim4IA 2025)
von: Schaer, Philipp, et al.
Veröffentlicht: (2025)
von: Schaer, Philipp, et al.
Veröffentlicht: (2025)
Scientometric Analysis of the German IR Community within TREC & CLEF
von: Kruff, A. K., et al.
Veröffentlicht: (2025)
von: Kruff, A. K., et al.
Veröffentlicht: (2025)
Report on the Workshop on Simulations for Information Access (Sim4IA 2024) at SIGIR 2024
von: Breuer, Timo, et al.
Veröffentlicht: (2024)
von: Breuer, Timo, et al.
Veröffentlicht: (2024)
Validating Search Query Simulations: A Taxonomy of Measures
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2026)
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2026)
Auditing Search Query Suggestion Bias Through Recursive Algorithm Interrogation
von: Haak, Fabian, et al.
Veröffentlicht: (2026)
von: Haak, Fabian, et al.
Veröffentlicht: (2026)
LISP -- A Rich Interaction Dataset and Loggable Interactive Search Platform
von: Friese, Jana Isabelle, et al.
Veröffentlicht: (2026)
von: Friese, Jana Isabelle, et al.
Veröffentlicht: (2026)
Cultural Analytics for Good: Building Inclusive Evaluation Frameworks for Historical IR
von: Datta, Suchana, et al.
Veröffentlicht: (2026)
von: Datta, Suchana, et al.
Veröffentlicht: (2026)
SPECTRA: Synthetic IR Test Collections with Relevance Oracles and Controlled Distractor Diagnostics
von: Liang, Eric
Veröffentlicht: (2026)
von: Liang, Eric
Veröffentlicht: (2026)
Synthetic Test Collections for Retrieval Evaluation
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
ASPIRE: Assistive System for Performance Evaluation in IR
von: Peikos, Georgios, et al.
Veröffentlicht: (2024)
von: Peikos, Georgios, et al.
Veröffentlicht: (2024)
A Comparison of Methods for Evaluating Generative IR
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2024)
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2024)
LLM-Driven Usefulness Labeling for IR Evaluation
von: Dewan, Mouly, et al.
Veröffentlicht: (2025)
von: Dewan, Mouly, et al.
Veröffentlicht: (2025)
Ranking Narrative Query Graphs for Biomedical Document Retrieval (Technical Report)
von: Kroll, Hermann, et al.
Veröffentlicht: (2024)
von: Kroll, Hermann, et al.
Veröffentlicht: (2024)
GenTREC: The First Test Collection Generated by Large Language Models for Evaluating Information Retrieval Systems
von: Türkmen, Mehmet Deniz, et al.
Veröffentlicht: (2025)
von: Türkmen, Mehmet Deniz, et al.
Veröffentlicht: (2025)
Building an Explainable Graph-based Biomedical Paper Recommendation System (Technical Report)
von: Kroll, Hermann, et al.
Veröffentlicht: (2024)
von: Kroll, Hermann, et al.
Veröffentlicht: (2024)
Improving the Reusability of Conversational Search Test Collections
von: Abbasiantaeb, Zahra, et al.
Veröffentlicht: (2025)
von: Abbasiantaeb, Zahra, et al.
Veröffentlicht: (2025)
Variations in Relevance Judgments and the Shelf Life of Test Collections
von: Parry, Andrew, et al.
Veröffentlicht: (2025)
von: Parry, Andrew, et al.
Veröffentlicht: (2025)
Investigating Bias in Political Search Query Suggestions by Relative Comparison with LLMs
von: Haak, Fabian, et al.
Veröffentlicht: (2024)
von: Haak, Fabian, et al.
Veröffentlicht: (2024)
DSEBench: A Test Collection for Explainable Dataset Search with Examples
von: Shi, Qing, et al.
Veröffentlicht: (2025)
von: Shi, Qing, et al.
Veröffentlicht: (2025)
Pairwise Comparison for Bias Identification and Quantification
von: Haak, Fabian, et al.
Veröffentlicht: (2025)
von: Haak, Fabian, et al.
Veröffentlicht: (2025)
STCALIR: Semi-Synthetic Test Collection for Algerian Legal Information Retrieval
von: Hatem, M'hamed Amine, et al.
Veröffentlicht: (2026)
von: Hatem, M'hamed Amine, et al.
Veröffentlicht: (2026)
SynDL: A Large-Scale Synthetic Test Collection for Passage Retrieval
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2024)
Can Generative LLMs Create Query Variants for Test Collections? An Exploratory Study
von: Alaofi, Marwah, et al.
Veröffentlicht: (2025)
von: Alaofi, Marwah, et al.
Veröffentlicht: (2025)
SimEval-IR: A Unified Toolkit and Benchmark Suite for Evaluating User Simulators and Search Sessions
von: Zerhoudi, Saber
Veröffentlicht: (2026)
von: Zerhoudi, Saber
Veröffentlicht: (2026)
Identifying Offline Metrics that Predict Online Impact: A Pragmatic Strategy for Real-World Recommender Systems
von: Wilm, Timo, et al.
Veröffentlicht: (2025)
von: Wilm, Timo, et al.
Veröffentlicht: (2025)
Make Any Collection Navigable: Methods for Constructing and Evaluating Hypergraph of Text
von: Alvarez, Dean E., et al.
Veröffentlicht: (2026)
von: Alvarez, Dean E., et al.
Veröffentlicht: (2026)
ExcluIR: Exclusionary Neural Information Retrieval
von: Zhang, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhang, Wenhao, et al.
Veröffentlicht: (2024)
Reproducing Adaptive Reranking for Reasoning-Intensive IR
von: Rathee, Mandeep, et al.
Veröffentlicht: (2026)
von: Rathee, Mandeep, et al.
Veröffentlicht: (2026)
Measuring Hypothesis Testing Errors in the Evaluation of Retrieval Systems
von: McKechnie, Jack, et al.
Veröffentlicht: (2025)
von: McKechnie, Jack, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Replicability Measures for Longitudinal Information Retrieval Evaluation
von: Keller, Jüri, et al.
Veröffentlicht: (2024) -
Evaluating Contrastive Feedback for Effective User Simulations
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025) -
Formalized Information Needs Improve Large-Language-Model Relevance Judgments
von: Keller, Jüri, et al.
Veröffentlicht: (2026) -
Context-Driven Interactive Query Simulations Based on Generative Large Language Models
von: Engelmann, Björn, et al.
Veröffentlicht: (2023) -
Sim4IA-Bench: A User Simulation Benchmark Suite for Next Query and Utterance Prediction
von: Kruff, Andreas Konstantin, et al.
Veröffentlicht: (2025)