An Empirical Study of Evaluating Long-form Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Xian, Ning, Fan, Yixing, Zhang, Ruqing, de Rijke, Maarten, Guo, Jiafeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Does Generative Retrieval Overcome the Limitations of Dense Retrieval?
by: Zhang, Yingchen, et al.
Published: (2025)
by: Zhang, Yingchen, et al.
Published: (2025)
From Relevance to Utility: Evidence Retrieval with Feedback for Fact Verification
by: Zhang, Hengran, et al.
Published: (2023)
by: Zhang, Hengran, et al.
Published: (2023)
Are Large Language Models Good at Utility Judgments?
by: Zhang, Hengran, et al.
Published: (2024)
by: Zhang, Hengran, et al.
Published: (2024)
On the Scaling of Robustness and Effectiveness in Dense Retrieval
by: Liu, Yu-An, et al.
Published: (2025)
by: Liu, Yu-An, et al.
Published: (2025)
Bootstrapped Pre-training with Dynamic Identifier Prediction for Generative Retrieval
by: Tang, Yubao, et al.
Published: (2024)
by: Tang, Yubao, et al.
Published: (2024)
Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-box Neural Ranking Models
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
Robust Information Retrieval
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
AdversarialCoT: Single-Document Retrieval Poisoning for LLM Reasoning
by: Song, Hongru, et al.
Published: (2026)
by: Song, Hongru, et al.
Published: (2026)
Continual Learning for Generative Retrieval over Dynamic Corpora
by: Chen, Jiangui, et al.
Published: (2023)
by: Chen, Jiangui, et al.
Published: (2023)
Multi-granular Adversarial Attacks against Black-box Neural Ranking Models
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
CorpusBrain++: A Continual Generative Pre-Training Framework for Knowledge-Intensive Language Tasks
by: Guo, Jiafeng, et al.
Published: (2024)
by: Guo, Jiafeng, et al.
Published: (2024)
Robust Neural Information Retrieval: An Adversarial and Out-of-distribution Perspective
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
On the Robustness of Generative Information Retrieval Models
by: Liu, Yu-An, et al.
Published: (2024)
by: Liu, Yu-An, et al.
Published: (2024)
Listwise Generative Retrieval Models via a Sequential Learning Process
by: Tang, Yubao, et al.
Published: (2024)
by: Tang, Yubao, et al.
Published: (2024)
Generative Retrieval Meets Multi-Graded Relevance
by: Tang, Yubao, et al.
Published: (2024)
by: Tang, Yubao, et al.
Published: (2024)
Chain-of-Thought Poisoning Attacks against R1-based Retrieval-Augmented Generation Systems
by: Song, Hongru, et al.
Published: (2025)
by: Song, Hongru, et al.
Published: (2025)
On the Capacity of Citation Generation by Large Language Models
by: Qian, Haosheng, et al.
Published: (2024)
by: Qian, Haosheng, et al.
Published: (2024)
Reason to Retrieve: Enhancing Query Understanding through Decomposition and Interpretation
by: Zhong, Yunfei, et al.
Published: (2025)
by: Zhong, Yunfei, et al.
Published: (2025)
The Silent Saboteur: Imperceptible Adversarial Attacks against Black-Box Retrieval-Augmented Generation Systems
by: Song, Hongru, et al.
Published: (2025)
by: Song, Hongru, et al.
Published: (2025)
Controlled Retrieval-augmented Context Evaluation for Long-form RAG
by: Ju, Jia-Huei, et al.
Published: (2025)
by: Ju, Jia-Huei, et al.
Published: (2025)
Generative Retrieval for Book search
by: Tang, Yubao, et al.
Published: (2025)
by: Tang, Yubao, et al.
Published: (2025)
TrustRAG: An Information Assistant with Retrieval Augmented Generation
by: Fan, Yixing, et al.
Published: (2025)
by: Fan, Yixing, et al.
Published: (2025)
VeriCite: Towards Reliable Citations in Retrieval-Augmented Generation via Rigorous Verification
by: Qian, Haosheng, et al.
Published: (2025)
by: Qian, Haosheng, et al.
Published: (2025)
LLMs as Sparse Retrievers:A Framework for First-Stage Product Search
by: Song, Hongru, et al.
Published: (2025)
by: Song, Hongru, et al.
Published: (2025)
Thinking Forward and Backward: Multi-Objective Reinforcement Learning for Retrieval-Augmented Reasoning
by: Wei, Wenda, et al.
Published: (2025)
by: Wei, Wenda, et al.
Published: (2025)
A Generative Framework for Personalized Sticker Retrieval
by: Zhou, Changjiang, et al.
Published: (2025)
by: Zhou, Changjiang, et al.
Published: (2025)
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
Retrieval-in-the-Chain: Bootstrapping Large Language Models for Generative Retrieval
by: Zhang, Yingchen, et al.
Published: (2025)
by: Zhang, Yingchen, et al.
Published: (2025)
Robust-IR @ SIGIR 2025: The First Workshop on Robust Information Retrieval
by: Liu, Yu-An, et al.
Published: (2025)
by: Liu, Yu-An, et al.
Published: (2025)
Rethinking the Evaluation of Dialogue Systems: Effects of User Feedback on Crowdworkers and LLMs
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
C2T-ID: Converting Semantic Codebooks to Textual Document Identifiers for Generative Search
by: Zhang, Yingchen, et al.
Published: (2025)
by: Zhang, Yingchen, et al.
Published: (2025)
Unidentified and Confounded? Understanding Two-Tower Models for Unbiased Learning to Rank (Extended Abstract)
by: Hager, Philipp, et al.
Published: (2025)
by: Hager, Philipp, et al.
Published: (2025)
Towards Two-Stage Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
ERASE -- A Real-World Aligned Benchmark for Unlearning in Recommender Systems
by: Lubitzsch, Pierre, et al.
Published: (2026)
by: Lubitzsch, Pierre, et al.
Published: (2026)
RankingSHAP -- Listwise Feature Attribution Explanations for Ranking Models
by: Heuss, Maria, et al.
Published: (2024)
by: Heuss, Maria, et al.
Published: (2024)
A First Look at Selection Bias in Preference Elicitation for Recommendation
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
HEC-GCN: Hypergraph Enhanced Cascading Graph Convolution Network for Multi-Behavior Recommendation
by: Yin, Yabo, et al.
Published: (2024)
by: Yin, Yabo, et al.
Published: (2024)
Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG
by: Bachyr, Omar El, et al.
Published: (2026)
by: Bachyr, Omar El, et al.
Published: (2026)
Cold-Starts in Generative Recommendation: A Reproducibility Study
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Do Images Clarify? A Study on the Effect of Images on Clarifying Questions in Conversational Search
by: Siro, Clemencia, et al.
Published: (2026)
by: Siro, Clemencia, et al.
Published: (2026)
Similar Items
-
Does Generative Retrieval Overcome the Limitations of Dense Retrieval?
by: Zhang, Yingchen, et al.
Published: (2025) -
From Relevance to Utility: Evidence Retrieval with Feedback for Fact Verification
by: Zhang, Hengran, et al.
Published: (2023) -
Are Large Language Models Good at Utility Judgments?
by: Zhang, Hengran, et al.
Published: (2024) -
On the Scaling of Robustness and Effectiveness in Dense Retrieval
by: Liu, Yu-An, et al.
Published: (2025) -
Bootstrapped Pre-training with Dynamic Identifier Prediction for Generative Retrieval
by: Tang, Yubao, et al.
Published: (2024)