CRS Arena: Crowdsourced Benchmarking of Conversational Recommender Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Bernard, Nolwenn, Joko, Hideaki, Hasibi, Faegheh, Balog, Krisztian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UserSimCRS v2: Simulation-Based Evaluation for Conversational Recommender Systems
by: Bernard, Nolwenn, et al.
Published: (2025)
by: Bernard, Nolwenn, et al.
Published: (2025)
FACE: A Fine-Grained Reference-Free Evaluator for Conversational Information Access
by: Joko, Hideaki, et al.
Published: (2025)
by: Joko, Hideaki, et al.
Published: (2025)
Identifying Breakdowns in Conversational Recommender Systems using User Simulation
by: Bernard, Nolwenn, et al.
Published: (2024)
by: Bernard, Nolwenn, et al.
Published: (2024)
Limitations of Current Evaluation Practices for Conversational Recommender Systems and the Potential of User Simulation
by: Bernard, Nolwenn, et al.
Published: (2025)
by: Bernard, Nolwenn, et al.
Published: (2025)
Towards a Formal Characterization of User Simulation Objectives in Conversational Information Access
by: Bernard, Nolwenn, et al.
Published: (2024)
by: Bernard, Nolwenn, et al.
Published: (2024)
WildClaims: Information Access Conversations in the Wild(Chat)
by: Joko, Hideaki, et al.
Published: (2025)
by: Joko, Hideaki, et al.
Published: (2025)
IAI MovieBot 2.0: An Enhanced Research Platform with Trainable Neural Components and Transparent User Modeling
by: Bernard, Nolwenn, et al.
Published: (2024)
by: Bernard, Nolwenn, et al.
Published: (2024)
SimLab: A Platform for Simulation-based Evaluation of Conversational Information Access Systems
by: Bernard, Nolwenn, et al.
Published: (2025)
by: Bernard, Nolwenn, et al.
Published: (2025)
Doing Personal LAPS: LLM-Augmented Dialogue Construction for Personalized Multi-Session Conversational Search
by: Joko, Hideaki, et al.
Published: (2024)
by: Joko, Hideaki, et al.
Published: (2024)
A Standardized Re-evaluation of Conversational Recommender Systems on the ReDial Dataset
by: Kostric, Ivica, et al.
Published: (2026)
by: Kostric, Ivica, et al.
Published: (2026)
Should We Tailor the Talk? Understanding the Impact of Conversational Styles on Preference Elicitation in Conversational Recommender Systems
by: Kostric, Ivica, et al.
Published: (2025)
by: Kostric, Ivica, et al.
Published: (2025)
Data Augmentation for Conversational AI
by: Soudani, Heydar, et al.
Published: (2023)
by: Soudani, Heydar, et al.
Published: (2023)
Generating Usage-related Questions for Preference Elicitation in Conversational Recommender Systems
by: Kostric, Ivica, et al.
Published: (2021)
by: Kostric, Ivica, et al.
Published: (2021)
SciNUP: Natural Language User Interest Profiles for Scientific Literature Recommendation
by: Arustashvili, Mariam, et al.
Published: (2025)
by: Arustashvili, Mariam, et al.
Published: (2025)
A Surprisingly Simple yet Effective Multi-Query Rewriting Method for Conversational Passage Retrieval
by: Kostric, Ivica, et al.
Published: (2024)
by: Kostric, Ivica, et al.
Published: (2024)
An Ecosystem for Personal Knowledge Graphs: A Survey and Research Roadmap
by: Skjæveland, Martin G., et al.
Published: (2023)
by: Skjæveland, Martin G., et al.
Published: (2023)
Know Your Users! Estimating User Domain Knowledge in Conversational Recommenders
by: Kostric, Ivica, et al.
Published: (2025)
by: Kostric, Ivica, et al.
Published: (2025)
Towards Reliable and Factual Response Generation: Detecting Unanswerable Questions in Information-Seeking Conversations
by: Łajewska, Weronika, et al.
Published: (2024)
by: Łajewska, Weronika, et al.
Published: (2024)
Why Uncertainty Estimation Methods Fall Short in RAG: An Axiomatic Analysis
by: Soudani, Heydar, et al.
Published: (2025)
by: Soudani, Heydar, et al.
Published: (2025)
Uncertainty Quantification for Retrieval-Augmented Reasoning
by: Soudani, Heydar, et al.
Published: (2025)
by: Soudani, Heydar, et al.
Published: (2025)
Dataset and Models for Item Recommendation Using Multi-Modal User Interactions
by: Bruun, Simone Borg, et al.
Published: (2024)
by: Bruun, Simone Borg, et al.
Published: (2024)
Trust Me on This: A User Study of Trustworthiness for RAG Responses
by: Łajewska, Weronika, et al.
Published: (2026)
by: Łajewska, Weronika, et al.
Published: (2026)
A Survey on Recent Advances in Conversational Data Generation
by: Soudani, Heydar, et al.
Published: (2024)
by: Soudani, Heydar, et al.
Published: (2024)
Towards Self-Contained Answers: Entity-Based Answer Rewriting in Conversational Search
by: Sekulić, Ivan, et al.
Published: (2024)
by: Sekulić, Ivan, et al.
Published: (2024)
Can Users Detect Biases or Factual Errors in Generated Responses in Conversational Information-Seeking?
by: Łajewska, Weronika, et al.
Published: (2024)
by: Łajewska, Weronika, et al.
Published: (2024)
Real World Conversational Entity Linking Requires More Than Zeroshots
by: Hoveyda, Mohanna, et al.
Published: (2024)
by: Hoveyda, Mohanna, et al.
Published: (2024)
LUMI: Unsupervised Intent Clustering with Multiple Pseudo-Labels
by: Lin, I-Fan, et al.
Published: (2025)
by: Lin, I-Fan, et al.
Published: (2025)
GINGER: Grounded Information Nugget-Based Generation of Responses
by: Łajewska, Weronika, et al.
Published: (2025)
by: Łajewska, Weronika, et al.
Published: (2025)
Graph-Embedding Empowered Entity Retrieval
by: Gerritse, Emma J., et al.
Published: (2025)
by: Gerritse, Emma J., et al.
Published: (2025)
Estimating the Usefulness of Clarifying Questions and Answers for Conversational Search
by: Sekulić, Ivan, et al.
Published: (2024)
by: Sekulić, Ivan, et al.
Published: (2024)
MemoCRS: Memory-enhanced Sequential Conversational Recommender Systems with Large Language Models
by: Xi, Yunjia, et al.
Published: (2024)
by: Xi, Yunjia, et al.
Published: (2024)
Explainability for Transparent Conversational Information-Seeking
by: Łajewska, Weronika, et al.
Published: (2024)
by: Łajewska, Weronika, et al.
Published: (2024)
Adaptive Orchestration of Modular Generative Information Access Systems
by: Hoveyda, Mohanna, et al.
Published: (2025)
by: Hoveyda, Mohanna, et al.
Published: (2025)
User Simulation for Evaluating Information Access Systems
by: Balog, Krisztian, et al.
Published: (2023)
by: Balog, Krisztian, et al.
Published: (2023)
Total Recall QA: A Verifiable Evaluation Suite for Deep Research Agents
by: Rafiee, Mahta, et al.
Published: (2026)
by: Rafiee, Mahta, et al.
Published: (2026)
Sim4IA-Bench: A User Simulation Benchmark Suite for Next Query and Utterance Prediction
by: Kruff, Andreas Konstantin, et al.
Published: (2025)
by: Kruff, Andreas Konstantin, et al.
Published: (2025)
User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation
by: Balog, Krisztian, et al.
Published: (2025)
by: Balog, Krisztian, et al.
Published: (2025)
SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems
by: Hao, Haochang, et al.
Published: (2026)
by: Hao, Haochang, et al.
Published: (2026)
Rankers, Judges, and Assistants: Towards Understanding the Interplay of LLMs in Information Retrieval Evaluation
by: Balog, Krisztian, et al.
Published: (2025)
by: Balog, Krisztian, et al.
Published: (2025)
Reproducing Complex Set-Compositional Information Retrieval
by: Degenhart, Vincent, et al.
Published: (2026)
by: Degenhart, Vincent, et al.
Published: (2026)
Similar Items
-
UserSimCRS v2: Simulation-Based Evaluation for Conversational Recommender Systems
by: Bernard, Nolwenn, et al.
Published: (2025) -
FACE: A Fine-Grained Reference-Free Evaluator for Conversational Information Access
by: Joko, Hideaki, et al.
Published: (2025) -
Identifying Breakdowns in Conversational Recommender Systems using User Simulation
by: Bernard, Nolwenn, et al.
Published: (2024) -
Limitations of Current Evaluation Practices for Conversational Recommender Systems and the Potential of User Simulation
by: Bernard, Nolwenn, et al.
Published: (2025) -
Towards a Formal Characterization of User Simulation Objectives in Conversational Information Access
by: Bernard, Nolwenn, et al.
Published: (2024)