Can Generative LLMs Create Query Variants for Test Collections? An Exploratory Study
Fuente:
arXiv
Saved in:
| Main Authors: | Alaofi, Marwah, Gallagher, Luke, Sanderson, Mark, Scholer, Falk, Thomas, Paul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Demographically-Inspired Query Variants Using an LLM
by: Alaofi, Marwah, et al.
Published: (2025)
by: Alaofi, Marwah, et al.
Published: (2025)
LLMs can be Fooled into Labelling a Document as Relevant (best café near me; this paper is perfectly relevant)
by: Alaofi, Marwah, et al.
Published: (2025)
by: Alaofi, Marwah, et al.
Published: (2025)
Generative Information Retrieval Evaluation
by: Alaofi, Marwah, et al.
Published: (2024)
by: Alaofi, Marwah, et al.
Published: (2024)
Online and Offline Evaluation in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)
by: Tavakoli, Leila, et al.
Published: (2024)
Personalisation of Generic Library Search Results Using Student Enrolment Information
by: Alaofi, Marwah, et al.
Published: (2015)
by: Alaofi, Marwah, et al.
Published: (2015)
Multi-stage Large Language Model Pipelines Can Outperform GPT-4o in Relevance Assessment
by: Schnabel, Julian A., et al.
Published: (2025)
by: Schnabel, Julian A., et al.
Published: (2025)
RMIT-ADM+S at the MMU-RAG NeurIPS 2025 Competition
by: Ran, Kun, et al.
Published: (2026)
by: Ran, Kun, et al.
Published: (2026)
Revisiting Query Variants: The Advantage of Retrieval Over Generation of Query Variants for Effective QPP
by: Tian, Fangzheng, et al.
Published: (2025)
by: Tian, Fangzheng, et al.
Published: (2025)
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
RAQG-QPP: Query Performance Prediction with Retrieved Query Variants and Retrieval Augmented Query Generation
by: Tian, Fangzheng, et al.
Published: (2026)
by: Tian, Fangzheng, et al.
Published: (2026)
Data Fusion of Synthetic Query Variants With Generative Large Language Models
by: Breuer, Timo
Published: (2024)
by: Breuer, Timo
Published: (2024)
Towards Investigating Biases in Spoken Conversational Search
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
Control Search Rankings, Control the World: What is a Good Search Engine?
by: Coghlan, Simon, et al.
Published: (2025)
by: Coghlan, Simon, et al.
Published: (2025)
The Effects of Demographic Instructions on LLM Personas
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2025)
by: de Paula, Angel Felipe Magnossão, et al.
Published: (2025)
QUIDS: Query Intent Description for Exploratory Search via Dual Space Modeling
by: Wang, Yumeng, et al.
Published: (2024)
by: Wang, Yumeng, et al.
Published: (2024)
SynDL: A Large-Scale Synthetic Test Collection for Passage Retrieval
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Diversity-Augmented Negative Sampling for Implicit Collaborative Filtering
by: Xuan, Yueqing, et al.
Published: (2025)
by: Xuan, Yueqing, et al.
Published: (2025)
Metamorphic Evaluation of ChatGPT as a Recommender System
by: Khirbat, Madhurima, et al.
Published: (2024)
by: Khirbat, Madhurima, et al.
Published: (2024)
Evaluating and Addressing Fairness Across User Groups in Negative Sampling for Recommender Systems
by: Xuan, Yueqing, et al.
Published: (2023)
by: Xuan, Yueqing, et al.
Published: (2023)
Augmented Knowledge Graph Querying leveraging LLMs
by: Arazzi, Marco, et al.
Published: (2025)
by: Arazzi, Marco, et al.
Published: (2025)
Generating Query Recommendations via LLMs
by: Bacciu, Andrea, et al.
Published: (2024)
by: Bacciu, Andrea, et al.
Published: (2024)
Stairway to Fairness: Connecting Group and Individual Fairness
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
A Reproducibility and Generalizability Study of Large Language Models for Query Generation
by: Staudinger, Moritz, et al.
Published: (2024)
by: Staudinger, Moritz, et al.
Published: (2024)
QueryExplorer: An Interactive Query Generation Assistant for Search and Exploration
by: Dhole, Kaustubh D., et al.
Published: (2024)
by: Dhole, Kaustubh D., et al.
Published: (2024)
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
by: Li, Minghan, et al.
Published: (2023)
by: Li, Minghan, et al.
Published: (2023)
Judging the Judges: A Collection of LLM-Generated Relevance Judgements
by: Rahmani, Hossein A., et al.
Published: (2025)
by: Rahmani, Hossein A., et al.
Published: (2025)
Walert: Putting Conversational Search Knowledge into Action by Building and Evaluating a Large Language Model-Powered Chatbot
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
Towards Detecting and Mitigating Cognitive Bias in Spoken Conversational Search
by: Ji, Kaixin, et al.
Published: (2024)
by: Ji, Kaixin, et al.
Published: (2024)
Improving the Reusability of Conversational Search Test Collections
by: Abbasiantaeb, Zahra, et al.
Published: (2025)
by: Abbasiantaeb, Zahra, et al.
Published: (2025)
Creating a Taxonomy for Retrieval Augmented Generation Applications
by: Nikishina, Irina, et al.
Published: (2024)
by: Nikishina, Irina, et al.
Published: (2024)
Evaluation of Temporal Change in IR Test Collections
by: Keller, Jüri, et al.
Published: (2024)
by: Keller, Jüri, et al.
Published: (2024)
Unsupervised Query Routing for Retrieval Augmented Generation
by: Mu, Feiteng, et al.
Published: (2025)
by: Mu, Feiteng, et al.
Published: (2025)
Generating Multi-Aspect Queries for Conversational Search
by: Abbasiantaeb, Zahra, et al.
Published: (2024)
by: Abbasiantaeb, Zahra, et al.
Published: (2024)
GenTREC: The First Test Collection Generated by Large Language Models for Evaluating Information Retrieval Systems
by: Türkmen, Mehmet Deniz, et al.
Published: (2025)
by: Türkmen, Mehmet Deniz, et al.
Published: (2025)
In a Few Words: Comparing Weak Supervision and LLMs for Short Query Intent Classification
by: Alexander, Daria, et al.
Published: (2025)
by: Alexander, Daria, et al.
Published: (2025)
Variations in Relevance Judgments and the Shelf Life of Test Collections
by: Parry, Andrew, et al.
Published: (2025)
by: Parry, Andrew, et al.
Published: (2025)
Combining Query Performance Predictors: A Reproducibility Study
by: Saha, Sourav, et al.
Published: (2025)
by: Saha, Sourav, et al.
Published: (2025)
Query Exposure Prediction for Groups of Documents in Rankings
by: Jaenich, Thomas, et al.
Published: (2024)
by: Jaenich, Thomas, et al.
Published: (2024)
LLMs Can Patch Up Missing Relevance Judgments in Evaluation
by: Upadhyay, Shivani, et al.
Published: (2024)
by: Upadhyay, Shivani, et al.
Published: (2024)
CTR-Guided Generative Query Suggestion in Conversational Search
by: Min, Erxue, et al.
Published: (2025)
by: Min, Erxue, et al.
Published: (2025)
Similar Items
-
Demographically-Inspired Query Variants Using an LLM
by: Alaofi, Marwah, et al.
Published: (2025) -
LLMs can be Fooled into Labelling a Document as Relevant (best café near me; this paper is perfectly relevant)
by: Alaofi, Marwah, et al.
Published: (2025) -
Generative Information Retrieval Evaluation
by: Alaofi, Marwah, et al.
Published: (2024) -
Online and Offline Evaluation in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024) -
Personalisation of Generic Library Search Results Using Student Enrolment Information
by: Alaofi, Marwah, et al.
Published: (2015)