When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Jiajie, Li, Yangchun, Chen, Guanyi, Fan, Rui, Bai, Xin, He, Tingting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Emotional Supporters often Use Multiple Strategies in a Single Turn
di: Bai, Xin, et al.
Pubblicazione: (2025)
di: Bai, Xin, et al.
Pubblicazione: (2025)
Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers
di: Heo, Chaewon, et al.
Pubblicazione: (2026)
di: Heo, Chaewon, et al.
Pubblicazione: (2026)
How Much Do LLMs Know About Chinese Zero Pronouns?
di: Li, Yifei, et al.
Pubblicazione: (2026)
di: Li, Yifei, et al.
Pubblicazione: (2026)
ESCoT: Towards Interpretable Emotional Support Dialogue Systems
di: Zhang, Tenggan, et al.
Pubblicazione: (2024)
di: Zhang, Tenggan, et al.
Pubblicazione: (2024)
Do Large Language Models Judge Error Severity Like Humans?
di: Sun, Diege, et al.
Pubblicazione: (2025)
di: Sun, Diege, et al.
Pubblicazione: (2025)
Emotional Support with LLM-based Empathetic Dialogue Generation
di: Wang, Shiquan, et al.
Pubblicazione: (2025)
di: Wang, Shiquan, et al.
Pubblicazione: (2025)
User-Aware Active Knowledge Acquisition for Emotional Support Dialogue
di: Xu, Mufan, et al.
Pubblicazione: (2026)
di: Xu, Mufan, et al.
Pubblicazione: (2026)
RecMind: Japanese Movie Recommendation Dialogue with Seeker's Internal State
di: Kodama, Takashi, et al.
Pubblicazione: (2024)
di: Kodama, Takashi, et al.
Pubblicazione: (2024)
On the Robustness of Knowledge Editing for Detoxification
di: Dong, Ming, et al.
Pubblicazione: (2026)
di: Dong, Ming, et al.
Pubblicazione: (2026)
How Do People Quantify Naturally: Evidence from Mandarin Picture Description
di: Zhang, Yayun, et al.
Pubblicazione: (2026)
di: Zhang, Yayun, et al.
Pubblicazione: (2026)
Dialogue Systems for Emotional Support via Value Reinforcement
di: Kim, Juhee, et al.
Pubblicazione: (2025)
di: Kim, Juhee, et al.
Pubblicazione: (2025)
IntentionESC: An Intention-Centered Framework for Enhancing Emotional Support in Dialogue Systems
di: Zhang, Xinjie, et al.
Pubblicazione: (2025)
di: Zhang, Xinjie, et al.
Pubblicazione: (2025)
What Do Humans Hear When Interacting? Experiments on Selective Listening for Evaluating ASR of Spoken Dialogue Systems
di: Mori, Kiyotada, et al.
Pubblicazione: (2025)
di: Mori, Kiyotada, et al.
Pubblicazione: (2025)
MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue
di: Zhang, Naifan, et al.
Pubblicazione: (2026)
di: Zhang, Naifan, et al.
Pubblicazione: (2026)
I Need Help! Evaluating LLM's Ability to Ask for Users' Support: A Case Study on Text-to-SQL Generation
di: Wu, Cheng-Kuang, et al.
Pubblicazione: (2024)
di: Wu, Cheng-Kuang, et al.
Pubblicazione: (2024)
TestAgent: Automatic Benchmarking and Exploratory Interaction for Evaluating LLMs in Vertical Domains
di: Wang, Wanying, et al.
Pubblicazione: (2024)
di: Wang, Wanying, et al.
Pubblicazione: (2024)
DiagESC: Dialogue Synthesis for Integrating Depression Diagnosis into Emotional Support Conversation
di: Seo, Seungyeon, et al.
Pubblicazione: (2024)
di: Seo, Seungyeon, et al.
Pubblicazione: (2024)
Intrinsic Task-based Evaluation for Referring Expression Generation
di: Chen, Guanyi, et al.
Pubblicazione: (2024)
di: Chen, Guanyi, et al.
Pubblicazione: (2024)
EmoDynamiX: Emotional Support Dialogue Strategy Prediction by Modelling MiXed Emotions and Discourse Dynamics
di: Wan, Chenwei, et al.
Pubblicazione: (2024)
di: Wan, Chenwei, et al.
Pubblicazione: (2024)
ThinkSwitcher: When to Think Hard, When to Think Fast
di: Liang, Guosheng, et al.
Pubblicazione: (2025)
di: Liang, Guosheng, et al.
Pubblicazione: (2025)
Prompt Optimization Is a Coin Flip: Diagnosing When It Helps in Compound AI Systems
di: Zhang, Xing, et al.
Pubblicazione: (2026)
di: Zhang, Xing, et al.
Pubblicazione: (2026)
Evaluating Social Bias in RAG Systems: When External Context Helps and Reasoning Hurts
di: Parihar, Shweta, et al.
Pubblicazione: (2026)
di: Parihar, Shweta, et al.
Pubblicazione: (2026)
CCNU at SemEval-2025 Task 3: Leveraging Internal and External Knowledge of Large Language Models for Multilingual Hallucination Annotation
di: Liu, Xu, et al.
Pubblicazione: (2025)
di: Liu, Xu, et al.
Pubblicazione: (2025)
HEART: A Unified Benchmark for Assessing Humans and LLMs in Emotional Support Dialogue
di: Iyer, Laya, et al.
Pubblicazione: (2026)
di: Iyer, Laya, et al.
Pubblicazione: (2026)
Scaling Worst-Case Optimal Datalog to GPUs
di: Sun, Yihao, et al.
Pubblicazione: (2026)
di: Sun, Yihao, et al.
Pubblicazione: (2026)
Certified Mitigation of Worst-Case LLM Copyright Infringement
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
Audio MultiChallenge: A Multi-Turn Evaluation of Spoken Dialogue Systems on Natural Human Interaction
di: Gosai, Advait, et al.
Pubblicazione: (2025)
di: Gosai, Advait, et al.
Pubblicazione: (2025)
Emotionally Intelligent Task-oriented Dialogue Systems: Architecture, Representation, and Optimisation
di: Feng, Shutong, et al.
Pubblicazione: (2025)
di: Feng, Shutong, et al.
Pubblicazione: (2025)
Infusing Emotions into Task-oriented Dialogue Systems: Understanding, Management, and Generation
di: Feng, Shutong, et al.
Pubblicazione: (2024)
di: Feng, Shutong, et al.
Pubblicazione: (2024)
RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems
di: Srinivasan, Adarsh, et al.
Pubblicazione: (2025)
di: Srinivasan, Adarsh, et al.
Pubblicazione: (2025)
"You are an expert annotator": Automatic Best-Worst-Scaling Annotations for Emotion Intensity Modeling
di: Bagdon, Christopher, et al.
Pubblicazione: (2024)
di: Bagdon, Christopher, et al.
Pubblicazione: (2024)
Does Collaborative Human-LM Dialogue Generation Help Information Extraction from Human Dialogues?
di: Lu, Bo-Ru, et al.
Pubblicazione: (2023)
di: Lu, Bo-Ru, et al.
Pubblicazione: (2023)
ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
di: Zhao, Haiquan, et al.
Pubblicazione: (2024)
di: Zhao, Haiquan, et al.
Pubblicazione: (2024)
MemEmo: Evaluating Emotion in Memory Systems of Agents
di: Liu, Peng, et al.
Pubblicazione: (2026)
di: Liu, Peng, et al.
Pubblicazione: (2026)
MMGeoLM: Hard Negative Contrastive Learning for Fine-Grained Geometric Understanding in Large Multimodal Models
di: Sun, Kai, et al.
Pubblicazione: (2025)
di: Sun, Kai, et al.
Pubblicazione: (2025)
Detecting Emotional Dynamic Trajectories: An Evaluation Framework for Emotional Support in Language Models
di: Tan, Zhouxing, et al.
Pubblicazione: (2025)
di: Tan, Zhouxing, et al.
Pubblicazione: (2025)
EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Spoken Dialogue Systems
di: Liu, Jingwen, et al.
Pubblicazione: (2025)
di: Liu, Jingwen, et al.
Pubblicazione: (2025)
EventWeave: A Dynamic Framework for Capturing Core and Supporting Events in Dialogue Systems
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
Think out Loud: Emotion Deducing Explanation in Dialogues
di: Li, Jiangnan, et al.
Pubblicazione: (2024)
di: Li, Jiangnan, et al.
Pubblicazione: (2024)
AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation
di: Wang, Ming, et al.
Pubblicazione: (2025)
di: Wang, Ming, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Emotional Supporters often Use Multiple Strategies in a Single Turn
di: Bai, Xin, et al.
Pubblicazione: (2025) -
Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers
di: Heo, Chaewon, et al.
Pubblicazione: (2026) -
How Much Do LLMs Know About Chinese Zero Pronouns?
di: Li, Yifei, et al.
Pubblicazione: (2026) -
ESCoT: Towards Interpretable Emotional Support Dialogue Systems
di: Zhang, Tenggan, et al.
Pubblicazione: (2024) -
Do Large Language Models Judge Error Severity Like Humans?
di: Sun, Diege, et al.
Pubblicazione: (2025)