Saved in:
| Main Authors: | Joko, Hideaki, Amirshahi, Shakiba, Clarke, Charles L. A., Hasibi, Faegheh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.17442 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FACE: A Fine-Grained Reference-Free Evaluator for Conversational Information Access
by: Joko, Hideaki, et al.
Published: (2025)
by: Joko, Hideaki, et al.
Published: (2025)
CRS Arena: Crowdsourced Benchmarking of Conversational Recommender Systems
by: Bernard, Nolwenn, et al.
Published: (2024)
by: Bernard, Nolwenn, et al.
Published: (2024)
Doing Personal LAPS: LLM-Augmented Dialogue Construction for Personalized Multi-Session Conversational Search
by: Joko, Hideaki, et al.
Published: (2024)
by: Joko, Hideaki, et al.
Published: (2024)
Data Augmentation for Conversational AI
by: Soudani, Heydar, et al.
Published: (2023)
by: Soudani, Heydar, et al.
Published: (2023)
Evaluating the Robustness of Retrieval-Augmented Generation to Adversarial Evidence in the Health Domain
by: Amirshahi, Shakiba, et al.
Published: (2025)
by: Amirshahi, Shakiba, et al.
Published: (2025)
Adaptive Orchestration of Modular Generative Information Access Systems
by: Hoveyda, Mohanna, et al.
Published: (2025)
by: Hoveyda, Mohanna, et al.
Published: (2025)
Why Uncertainty Estimation Methods Fall Short in RAG: An Axiomatic Analysis
by: Soudani, Heydar, et al.
Published: (2025)
by: Soudani, Heydar, et al.
Published: (2025)
Uncertainty Quantification for Retrieval-Augmented Reasoning
by: Soudani, Heydar, et al.
Published: (2025)
by: Soudani, Heydar, et al.
Published: (2025)
A Survey on Recent Advances in Conversational Data Generation
by: Soudani, Heydar, et al.
Published: (2024)
by: Soudani, Heydar, et al.
Published: (2024)
Real World Conversational Entity Linking Requires More Than Zeroshots
by: Hoveyda, Mohanna, et al.
Published: (2024)
by: Hoveyda, Mohanna, et al.
Published: (2024)
LUMI: Unsupervised Intent Clustering with Multiple Pseudo-Labels
by: Lin, I-Fan, et al.
Published: (2025)
by: Lin, I-Fan, et al.
Published: (2025)
Graph-Embedding Empowered Entity Retrieval
by: Gerritse, Emma J., et al.
Published: (2025)
by: Gerritse, Emma J., et al.
Published: (2025)
Reproducing Complex Set-Compositional Information Retrieval
by: Degenhart, Vincent, et al.
Published: (2026)
by: Degenhart, Vincent, et al.
Published: (2026)
Total Recall QA: A Verifiable Evaluation Suite for Deep Research Agents
by: Rafiee, Mahta, et al.
Published: (2026)
by: Rafiee, Mahta, et al.
Published: (2026)
LLM-Evaluation Tropes: Perspectives on the Validity of LLM-Evaluations
by: Dietz, Laura, et al.
Published: (2025)
by: Dietz, Laura, et al.
Published: (2025)
OrLog: Resolving Complex Queries with LLMs and Probabilistic Reasoning
by: Hoveyda, Mohanna, et al.
Published: (2026)
by: Hoveyda, Mohanna, et al.
Published: (2026)
WildVis: Open Source Visualizer for Million-Scale Chat Logs in the Wild
by: Deng, Yuntian, et al.
Published: (2024)
by: Deng, Yuntian, et al.
Published: (2024)
Fréchet Distance for Offline Evaluation of Information Retrieval Systems with Sparse Labels
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
Annotative Indexing
by: Clarke, Charles L. A.
Published: (2024)
by: Clarke, Charles L. A.
Published: (2024)
Generative Information Retrieval Evaluation
by: Alaofi, Marwah, et al.
Published: (2024)
by: Alaofi, Marwah, et al.
Published: (2024)
A Comparison of Methods for Evaluating Generative IR
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
Making Sense of Data in the Wild: Data Analysis Automation at Scale
by: Graziani, Mara, et al.
Published: (2025)
by: Graziani, Mara, et al.
Published: (2025)
Adapting Standard Retrieval Benchmarks to Evaluate Generated Answers
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
Towards a Formal Characterization of User Simulation Objectives in Conversational Information Access
by: Bernard, Nolwenn, et al.
Published: (2024)
by: Bernard, Nolwenn, et al.
Published: (2024)
SimLab: A Platform for Simulation-based Evaluation of Conversational Information Access Systems
by: Bernard, Nolwenn, et al.
Published: (2025)
by: Bernard, Nolwenn, et al.
Published: (2025)
Benchmarking LLM-based Relevance Judgment Methods
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance Judgment
by: Arabzadeh, Negar, et al.
Published: (2025)
by: Arabzadeh, Negar, et al.
Published: (2025)
LLM-based relevance assessment still can't replace human relevance assessment
by: Clarke, Charles L. A., et al.
Published: (2024)
by: Clarke, Charles L. A., et al.
Published: (2024)
EMPRA: Embedding Perturbation Rank Attack against Neural Ranking Models
by: Bigdeli, Amin, et al.
Published: (2024)
by: Bigdeli, Amin, et al.
Published: (2024)
Efficient Vector Search in the Wild: One Model for Multi-K Queries
by: Peng, Yifan, et al.
Published: (2026)
by: Peng, Yifan, et al.
Published: (2026)
Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness
by: Loweimi, Erfan, et al.
Published: (2025)
by: Loweimi, Erfan, et al.
Published: (2025)
ChatShopBuddy: Towards Reliable Conversational Shopping Agents via Reinforcement Learning
by: Cheng, Yiruo, et al.
Published: (2026)
by: Cheng, Yiruo, et al.
Published: (2026)
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
by: Wu, Bin, et al.
Published: (2026)
by: Wu, Bin, et al.
Published: (2026)
Public Profile Matters: A Scalable Integrated Approach to Recommend Citations in the Wild
by: Goyal, Karan, et al.
Published: (2026)
by: Goyal, Karan, et al.
Published: (2026)
Resources for Automated Evaluation of Assistive RAG Systems that Help Readers with News Trustworthiness Assessment
by: Zhang, Dake, et al.
Published: (2026)
by: Zhang, Dake, et al.
Published: (2026)
Snap and Diagnose: An Advanced Multimodal Retrieval System for Identifying Plant Diseases in the Wild
by: Wei, Tianqi, et al.
Published: (2024)
by: Wei, Tianqi, et al.
Published: (2024)
Report on the 1st Workshop on Large Language Model for Evaluation in Information Retrieval (LLM4Eval 2024) at SIGIR 2024
by: Rahmani, Hossein A., et al.
Published: (2024)
by: Rahmani, Hossein A., et al.
Published: (2024)
Access to Library Information Resources by University Students during COVID-19 Pandemic in Africa: A Systematic Literature Review
by: Shikali, Joyce Charles, et al.
Published: (2024)
by: Shikali, Joyce Charles, et al.
Published: (2024)
Towards Context-Aware Adaptation in Extended Reality: A Design Space for XR Interfaces and an Adaptive Placement Strategy
by: Davari, Shakiba, et al.
Published: (2024)
by: Davari, Shakiba, et al.
Published: (2024)
Agent-centric Information Access
by: Kanoulas, Evangelos, et al.
Published: (2025)
by: Kanoulas, Evangelos, et al.
Published: (2025)
Similar Items
-
FACE: A Fine-Grained Reference-Free Evaluator for Conversational Information Access
by: Joko, Hideaki, et al.
Published: (2025) -
CRS Arena: Crowdsourced Benchmarking of Conversational Recommender Systems
by: Bernard, Nolwenn, et al.
Published: (2024) -
Doing Personal LAPS: LLM-Augmented Dialogue Construction for Personalized Multi-Session Conversational Search
by: Joko, Hideaki, et al.
Published: (2024) -
Data Augmentation for Conversational AI
by: Soudani, Heydar, et al.
Published: (2023) -
Evaluating the Robustness of Retrieval-Augmented Generation to Adversarial Evidence in the Health Domain
by: Amirshahi, Shakiba, et al.
Published: (2025)