Dagstuhl Perspectives Workshop 24352 -- Conversational Agents: A Framework for Evaluation (CAFE): Manifesto
Fuente:
arXiv
Saved in:
| Main Authors: | Bauer, Christine, Chen, Li, Ferro, Nicola, Fuhr, Norbert, Anand, Avishek, Breuer, Timo, Faggioli, Guglielmo, Frieder, Ophir, Joho, Hideo, Karlgren, Jussi, Kiesel, Johannes, Knijnenburg, Bart P., Lipani, Aldo, Michiels, Lien, Papenmeier, Andrea, Pera, Maria Soledad, Sanderson, Mark, Sanner, Scott, Stein, Benno, Trippas, Johanne R., Verspoor, Karin, Willemsen, Martijn C |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Manifesto from Dagstuhl Perspectives Workshop 24452 -- Reframing Technical Debt
by: Avgeriou, Paris, et al.
Published: (2025)
by: Avgeriou, Paris, et al.
Published: (2025)
Online and Offline Evaluation in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)
by: Tavakoli, Leila, et al.
Published: (2024)
Are We Wasting Time? A Fast, Accurate Performance Evaluation Framework for Knowledge Graph Link Predictors
by: Cornell, Filip, et al.
Published: (2024)
by: Cornell, Filip, et al.
Published: (2024)
Mario at EXIST 2025: A Simple Gateway to Effective Multilingual Sexism Detection
by: Tian, Lin, et al.
Published: (2025)
by: Tian, Lin, et al.
Published: (2025)
Can Users Detect Biases or Factual Errors in Generated Responses in Conversational Information-Seeking?
by: Łajewska, Weronika, et al.
Published: (2024)
by: Łajewska, Weronika, et al.
Published: (2024)
Explainability for Transparent Conversational Information-Seeking
by: Łajewska, Weronika, et al.
Published: (2024)
by: Łajewska, Weronika, et al.
Published: (2024)
Roadmap for Edge AI: A Dagstuhl Perspective
by: Ding, Aaron Yi, et al.
Published: (2021)
by: Ding, Aaron Yi, et al.
Published: (2021)
AMNESIA: A Large Scale Medical Unlearning Benchmark Suite with Disease-Informed Analysis
by: Davoudi, Saeedeh, et al.
Published: (2026)
by: Davoudi, Saeedeh, et al.
Published: (2026)
Towards Investigating Biases in Spoken Conversational Search
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
by: Cherumanal, Sachin Pathiyan, et al.
Published: (2024)
Multi-stage Large Language Model Pipelines Can Outperform GPT-4o in Relevance Assessment
by: Schnabel, Julian A., et al.
Published: (2025)
by: Schnabel, Julian A., et al.
Published: (2025)
A chance to learn : knowledge and finance for education in Sub-Saharan Africa / Adriaan Verspoor
by: Verspoor, Adriaan
Published: (2001)
by: Verspoor, Adriaan
Published: (2001)
Pathways two change : improving the quality of education in developing countries / Adriaan Verspoor
by: Verspoor, Adriaan
Published: (1989)
by: Verspoor, Adriaan
Published: (1989)
At the crossroads : choices for secondary education in Sub-Saharan Africa / Adriaan Verspoor
by: Verspoor, Adriaan
by: Verspoor, Adriaan
Challenges to the planning of education / Adriaan Verspoor
by: Verspoor, Adriaan
Published: (1992)
by: Verspoor, Adriaan
Published: (1992)
Disambiguating Complexity: From CAF to CAFIC: A Commentary on “Complexity and Difficulty in Second Language Acquisition: A Theoretical and Methodological Overview”
by: Marjolijn Verspoor
Published: (2024)
by: Marjolijn Verspoor
Published: (2024)
Genetic Approach to Mitigate Hallucination in Generative IR
by: Kulkarni, Hrishikesh, et al.
Published: (2024)
by: Kulkarni, Hrishikesh, et al.
Published: (2024)
LexBoost: Improving Lexical Document Retrieval with Nearest Neighbors
by: Kulkarni, Hrishikesh, et al.
Published: (2024)
by: Kulkarni, Hrishikesh, et al.
Published: (2024)
Modern Socio-Technical Perspectives on Privacy
by: Bart P. Knijnenburg
by: Bart P. Knijnenburg
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
TARAZ: Persian Short-Answer Question Benchmark for Cultural Evaluation of Language Models
by: Iranmanesh, Reihaneh, et al.
Published: (2026)
by: Iranmanesh, Reihaneh, et al.
Published: (2026)
Pore water geochemistry of sediment core GeoB24352-1
by: Zabel, Matthias, et al.
Published: (2023)
by: Zabel, Matthias, et al.
Published: (2023)
Can We Hide Machines in the Crowd? Quantifying Equivalence in LLM-in-the-loop Annotation Tasks
by: He, Jiaman, et al.
Published: (2025)
by: He, Jiaman, et al.
Published: (2025)
Characterising Topic Familiarity and Query Specificity Using Eye-Tracking Data
by: He, Jiaman, et al.
Published: (2025)
by: He, Jiaman, et al.
Published: (2025)
Does Cognitive Load Affect Human Accuracy in Detecting Voice-Based Deepfakes?
by: Gohsen, Marcel, et al.
Published: (2026)
by: Gohsen, Marcel, et al.
Published: (2026)
10 | THE ROLE OF ARTIFICIAL INTELLIGENCE IN CLINICAL TRIAL DESIGN AND ANALYSIS
by: S. Michiels
Published: (2025)
by: S. Michiels
Published: (2025)
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
by: Shi, Zhengxiang, et al.
Published: (2023)
by: Shi, Zhengxiang, et al.
Published: (2023)
GRIT: Graph-based Recall Improvement for Task-oriented E-commerce Queries
by: Kulkarni, Hrishikesh, et al.
Published: (2025)
by: Kulkarni, Hrishikesh, et al.
Published: (2025)
Intercept Cancer: Cancer Pre-Screening with Large Scale Healthcare Foundation Models
by: Sun, Liwen, et al.
Published: (2025)
by: Sun, Liwen, et al.
Published: (2025)
A Picture of Agentic Search
by: Pezzuti, Francesca, et al.
Published: (2026)
by: Pezzuti, Francesca, et al.
Published: (2026)
Understanding Modality Preferences in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)
by: Tavakoli, Leila, et al.
Published: (2024)
Covington, Coline. Who's to Blame? Collective Guilt on Trial. Routledge. 2023. Pp. 172. Pbk. £18.99.
by: Hessel Willemsen
Published: (2024)
by: Hessel Willemsen
Published: (2024)
Characterizing Personality from Eye-Tracking: The Role of Gaze and Its Absence in Interactive Search Environments
by: He, Jiaman, et al.
Published: (2026)
by: He, Jiaman, et al.
Published: (2026)
DRAMA: Domain Retrieval using Adaptive Module Allocation
by: Kasela, Pranav, et al.
Published: (2026)
by: Kasela, Pranav, et al.
Published: (2026)
EMBRE: Entity-aware Masking for Biomedical Relation Extraction
by: Li, Mingjie, et al.
Published: (2024)
by: Li, Mingjie, et al.
Published: (2024)
A simulation-based approach to the fluid-structure interaction inside fatigue cracks in hydraulic components
by: Michiels, Lukas
Published: (2026)
by: Michiels, Lukas
Published: (2026)
A simulation-based approach to the fluid-structure interaction inside fatigue cracks in hydraulic components
by: Michiels, Lukas
Published: (2026)
by: Michiels, Lukas
Published: (2026)
JUSTICIA Y DERECHO DESDE LA PERSPECTIVA FILOSÓFICA DEL ORDEN SOCIAL Y CULTURA JURÍDICA
by: Alizia Agnelli Faggioli
Published: (2019)
by: Alizia Agnelli Faggioli
Published: (2019)
INTERPRETACIÓN DE LA TUTELA JUDICIAL A LA NATURALEZA COMO GARANTÍA DE JUSTICIA EN ECUADOR
by: Alizia Agnelli Faggioli
Published: (2018)
by: Alizia Agnelli Faggioli
Published: (2018)
LAS TECNOLOGÍAS DE LA INFORMACIÓN Y LA COMUNICACIÓN Y SU AVANCE EN EL CONTEXTO EDUCATIVO
by: Alizia Agnelli Faggioli
Published: (2020)
by: Alizia Agnelli Faggioli
Published: (2020)
VINCULACIÓN HECHO Y DEBER SOCIAL TRABAJO CON EL PRINCIPIO PROTECTOR
by: Alizia Agnelli Faggioli
Published: (2020)
by: Alizia Agnelli Faggioli
Published: (2020)
Similar Items
-
Manifesto from Dagstuhl Perspectives Workshop 24452 -- Reframing Technical Debt
by: Avgeriou, Paris, et al.
Published: (2025) -
Online and Offline Evaluation in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024) -
Are We Wasting Time? A Fast, Accurate Performance Evaluation Framework for Knowledge Graph Link Predictors
by: Cornell, Filip, et al.
Published: (2024) -
Mario at EXIST 2025: A Simple Gateway to Effective Multilingual Sexism Detection
by: Tian, Lin, et al.
Published: (2025) -
Can Users Detect Biases or Factual Errors in Generated Responses in Conversational Information-Seeking?
by: Łajewska, Weronika, et al.
Published: (2024)