Socio-Emotional Response Generation: A Human Evaluation Protocol for LLM-Based Conversational Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vanel, Lorraine, Vela, Ariel R. Ramos, Yacoubi, Alya, Clavel, Chloé |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Think Twice: A Human-like Two-stage Conversational Agent for Emotional Response Generation
von: Qian, Yushan, et al.
Veröffentlicht: (2023)
von: Qian, Yushan, et al.
Veröffentlicht: (2023)
Towards Human-centered Proactive Conversational Agents
von: Deng, Yang, et al.
Veröffentlicht: (2024)
von: Deng, Yang, et al.
Veröffentlicht: (2024)
EmoXpt: Analyzing Emotional Variances in Human Comments and LLM-Generated Responses
von: Pyreddy, Shireesh Reddy, et al.
Veröffentlicht: (2025)
von: Pyreddy, Shireesh Reddy, et al.
Veröffentlicht: (2025)
Toward Safe and Human-Aligned Game Conversational Recommendation via Multi-Agent Decomposition
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
Dagstuhl Perspectives Workshop 24352 -- Conversational Agents: A Framework for Evaluation (CAFE): Manifesto
von: Bauer, Christine, et al.
Veröffentlicht: (2025)
von: Bauer, Christine, et al.
Veröffentlicht: (2025)
Telephone Surveys Meet Conversational AI: Evaluating a LLM-Based Telephone Survey System at Scale
von: Lang, Max M., et al.
Veröffentlicht: (2025)
von: Lang, Max M., et al.
Veröffentlicht: (2025)
LLM Confidence Evaluation Measures in Zero-Shot CSS Classification
von: Farr, David, et al.
Veröffentlicht: (2024)
von: Farr, David, et al.
Veröffentlicht: (2024)
Human-LLM Collaborative Construction of a Cantonese Emotion Lexicon
von: Zhang, Yusong, et al.
Veröffentlicht: (2024)
von: Zhang, Yusong, et al.
Veröffentlicht: (2024)
Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking
von: Roitero, Kevin, et al.
Veröffentlicht: (2025)
von: Roitero, Kevin, et al.
Veröffentlicht: (2025)
Context Does Matter: Implications for Crowdsourced Evaluation Labels in Task-Oriented Dialogue Systems
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
Emotionally Aware Moderation: The Potential of Emotion Monitoring in Shaping Healthier Social Media Conversations
von: Su, Xiaotian, et al.
Veröffentlicht: (2025)
von: Su, Xiaotian, et al.
Veröffentlicht: (2025)
Do Images Clarify? A Study on the Effect of Images on Clarifying Questions in Conversational Search
von: Siro, Clemencia, et al.
Veröffentlicht: (2026)
von: Siro, Clemencia, et al.
Veröffentlicht: (2026)
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
Unraveling Code-Mixing Patterns in Migration Discourse: Automated Detection and Analysis of Online Conversations on Reddit
von: Vitiugin, Fedor, et al.
Veröffentlicht: (2024)
von: Vitiugin, Fedor, et al.
Veröffentlicht: (2024)
Serendipity with Generative AI: Repurposing knowledge components during polycrisis with a Viable Systems Model approach
von: Fletcher, Gordon, et al.
Veröffentlicht: (2025)
von: Fletcher, Gordon, et al.
Veröffentlicht: (2025)
MEGAnno+: A Human-LLM Collaborative Annotation System
von: Kim, Hannah, et al.
Veröffentlicht: (2024)
von: Kim, Hannah, et al.
Veröffentlicht: (2024)
EHR-MCP: Real-world Evaluation of Clinical Information Retrieval by Large Language Models via Model Context Protocol
von: Masayoshi, Kanato, et al.
Veröffentlicht: (2025)
von: Masayoshi, Kanato, et al.
Veröffentlicht: (2025)
Conversations Gone Awry, But Then? Evaluating Conversational Forecasting Models
von: Tran, Son Quoc, et al.
Veröffentlicht: (2025)
von: Tran, Son Quoc, et al.
Veröffentlicht: (2025)
Verify as You Go: An LLM-Powered Browser Extension for Fake News Detection
von: Sallami, Dorsaf, et al.
Veröffentlicht: (2026)
von: Sallami, Dorsaf, et al.
Veröffentlicht: (2026)
Tip of the Tongue Query Elicitation for Simulated Evaluation
von: He, Yifan, et al.
Veröffentlicht: (2025)
von: He, Yifan, et al.
Veröffentlicht: (2025)
What should I wear to a party in a Greek taverna? Evaluation for Conversational Agents in the Fashion Domain
von: Maronikolakis, Antonis, et al.
Veröffentlicht: (2024)
von: Maronikolakis, Antonis, et al.
Veröffentlicht: (2024)
Human and LLM Biases in Hate Speech Annotations: A Socio-Demographic Analysis of Annotators and Targets
von: Giorgi, Tommaso, et al.
Veröffentlicht: (2024)
von: Giorgi, Tommaso, et al.
Veröffentlicht: (2024)
A Survey on LLM-based Conversational User Simulation
von: Ni, Bo, et al.
Veröffentlicht: (2026)
von: Ni, Bo, et al.
Veröffentlicht: (2026)
Pearl: Personalizing Large Language Model Writing Assistants with Generation-Calibrated Retrievers
von: Mysore, Sheshera, et al.
Veröffentlicht: (2023)
von: Mysore, Sheshera, et al.
Veröffentlicht: (2023)
Causal Autoencoder-like Generation of Feedback Fuzzy Cognitive Maps with an LLM Agent
von: Panda, Akash Kumar, et al.
Veröffentlicht: (2025)
von: Panda, Akash Kumar, et al.
Veröffentlicht: (2025)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
von: Saha, Koustuv, et al.
Veröffentlicht: (2025)
Evolving Paradigms in Task-Based Search and Learning: A Comparative Analysis of Traditional Search Engine with LLM-Enhanced Conversational Search System
von: Guan, Zhitong, et al.
Veröffentlicht: (2025)
von: Guan, Zhitong, et al.
Veröffentlicht: (2025)
Rewriting Conversational Utterances with Instructed Large Language Models
von: Galimzhanova, Elnara, et al.
Veröffentlicht: (2024)
von: Galimzhanova, Elnara, et al.
Veröffentlicht: (2024)
Human Latency Conversational Turns for Spoken Avatar Systems
von: Jacoby, Derek, et al.
Veröffentlicht: (2024)
von: Jacoby, Derek, et al.
Veröffentlicht: (2024)
EICAP: Deep Dive in Assessment and Enhancement of Large Language Models in Emotional Intelligence through Multi-Turn Conversations
von: Nazar, Nizi, et al.
Veröffentlicht: (2025)
von: Nazar, Nizi, et al.
Veröffentlicht: (2025)
KrishokBondhu: A Retrieval-Augmented Voice-Based Agricultural Advisory Call Center for Bengali Farmers
von: Ameen, Mohd Ruhul, et al.
Veröffentlicht: (2025)
von: Ameen, Mohd Ruhul, et al.
Veröffentlicht: (2025)
Conversations in Space: Structuring Non-Linear LLM Interactions on a Canvas
von: Amin, Rifat Mehreen, et al.
Veröffentlicht: (2026)
von: Amin, Rifat Mehreen, et al.
Veröffentlicht: (2026)
From PARIS to LE-PARIS: Toward Patent Response Automation with Recommender Systems and Collaborative Large Language Models
von: Chu, Jung-Mei, et al.
Veröffentlicht: (2024)
von: Chu, Jung-Mei, et al.
Veröffentlicht: (2024)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
von: Huang, Dun-Ming, et al.
Veröffentlicht: (2024)
von: Huang, Dun-Ming, et al.
Veröffentlicht: (2024)
Living the Novel: A System for Generating Self-Training Timeline-Aware Conversational Agents from Novels
von: Huang, Yifei, et al.
Veröffentlicht: (2025)
von: Huang, Yifei, et al.
Veröffentlicht: (2025)
Evaluating LLM-Generated Q&A Test: a Student-Centered Study
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
From Chatbots to Confidants: A Cross-Cultural Study of LLM Adoption for Emotional Support
von: Amat-Lefort, Natalia, et al.
Veröffentlicht: (2026)
von: Amat-Lefort, Natalia, et al.
Veröffentlicht: (2026)
Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications
von: Tavakoli, Leila, et al.
Veröffentlicht: (2025)
von: Tavakoli, Leila, et al.
Veröffentlicht: (2025)
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Exploring Gender Biases in Language Patterns of Human-Conversational Agent Conversations
von: Liu, Weizi
Veröffentlicht: (2024)
von: Liu, Weizi
Veröffentlicht: (2024)
Ähnliche Einträge
-
Think Twice: A Human-like Two-stage Conversational Agent for Emotional Response Generation
von: Qian, Yushan, et al.
Veröffentlicht: (2023) -
Towards Human-centered Proactive Conversational Agents
von: Deng, Yang, et al.
Veröffentlicht: (2024) -
EmoXpt: Analyzing Emotional Variances in Human Comments and LLM-Generated Responses
von: Pyreddy, Shireesh Reddy, et al.
Veröffentlicht: (2025) -
Toward Safe and Human-Aligned Game Conversational Recommendation via Multi-Agent Decomposition
von: Hui, Zheng, et al.
Veröffentlicht: (2025) -
Dagstuhl Perspectives Workshop 24352 -- Conversational Agents: A Framework for Evaluation (CAFE): Manifesto
von: Bauer, Christine, et al.
Veröffentlicht: (2025)