ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Haiquan, Li, Lingyu, Chen, Shisong, Kong, Shuqi, Wang, Jiaan, Huang, Kexin, Gu, Tianle, Wang, Yixu, Jian, Wang, Liang, Dandan, Li, Zhixu, Teng, Yan, Xiao, Yanghua, Wang, Yingchun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
von: Li, Lingyu, et al.
Veröffentlicht: (2024)
von: Li, Lingyu, et al.
Veröffentlicht: (2024)
OVEL: Large Language Model as Memory Manager for Online Video Entity Linking
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
HoneypotNet: Backdoor Attacks Against Model Extraction
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
The Other Mind: How Language Models Exhibit Human Temporal Cognition
von: Li, Lingyu, et al.
Veröffentlicht: (2025)
von: Li, Lingyu, et al.
Veröffentlicht: (2025)
M^2ConceptBase: A Fine-Grained Aligned Concept-Centric Multimodal Knowledge Base
von: Zha, Zhiwei, et al.
Veröffentlicht: (2023)
von: Zha, Zhiwei, et al.
Veröffentlicht: (2023)
LinguaSafe: A Comprehensive Multilingual Safety Benchmark for Large Language Models
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
Probing the Robustness of Large Language Models Safety to Latent Perturbations
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations
von: Zhu, Jie, et al.
Veröffentlicht: (2026)
von: Zhu, Jie, et al.
Veröffentlicht: (2026)
Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes?
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
Mechanistic Origin of Moral Indifference in Language Models
von: Li, Lingyu, et al.
Veröffentlicht: (2026)
von: Li, Lingyu, et al.
Veröffentlicht: (2026)
From Sparse Decisions to Dense Reasoning: A Multi-attribute Trajectory Paradigm for Multimodal Moderation
von: Gu, Tianle, et al.
Veröffentlicht: (2026)
von: Gu, Tianle, et al.
Veröffentlicht: (2026)
Improving the Robustness of Knowledge-Grounded Dialogue via Contrastive Learning
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
StolenLoRA: Exploring LoRA Extraction Attacks via Synthetic Data
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
Dr. Bench: A Multidimensional Evaluation for Deep Research Agents, from Answers to Reports
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
OpenRT: An Open-Source Red Teaming Framework for Multimodal LLMs
von: Wang, Xin, et al.
Veröffentlicht: (2026)
von: Wang, Xin, et al.
Veröffentlicht: (2026)
ESC-Judge: A Framework for Comparing Emotional Support Conversational Agents
von: Madani, Navid, et al.
Veröffentlicht: (2025)
von: Madani, Navid, et al.
Veröffentlicht: (2025)
CauESC: A Causal Aware Model for Emotional Support Conversation
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
MEOW: MEMOry Supervised LLM Unlearning Via Inverted Facts
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
IntentionESC: An Intention-Centered Framework for Enhancing Emotional Support in Dialogue Systems
von: Zhang, Xinjie, et al.
Veröffentlicht: (2025)
von: Zhang, Xinjie, et al.
Veröffentlicht: (2025)
DiagESC: Dialogue Synthesis for Integrating Depression Diagnosis into Emotional Support Conversation
von: Seo, Seungyeon, et al.
Veröffentlicht: (2024)
von: Seo, Seungyeon, et al.
Veröffentlicht: (2024)
SentGuard: Sentence-Level Streaming Guardrails for Large Language Models
von: Yu, Jiaqi, et al.
Veröffentlicht: (2026)
von: Yu, Jiaqi, et al.
Veröffentlicht: (2026)
A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment
von: Jiang, Sihang, et al.
Veröffentlicht: (2026)
von: Jiang, Sihang, et al.
Veröffentlicht: (2026)
Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs
von: Chen, Yunhao, et al.
Veröffentlicht: (2025)
von: Chen, Yunhao, et al.
Veröffentlicht: (2025)
JailBound: Jailbreaking Internal Safety Boundaries of Vision-Language Models
von: Song, Jiaxin, et al.
Veröffentlicht: (2025)
von: Song, Jiaxin, et al.
Veröffentlicht: (2025)
Adaptive Ordered Information Extraction with Deep Reinforcement Learning
von: Huang, Wenhao, et al.
Veröffentlicht: (2023)
von: Huang, Wenhao, et al.
Veröffentlicht: (2023)
Is There a One-Model-Fits-All Approach to Information Extraction? Revisiting Task Definition Biases
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
Towards Context-Invariant Safety Alignment for Large Language Models
von: Wang, Yixu, et al.
Veröffentlicht: (2026)
von: Wang, Yixu, et al.
Veröffentlicht: (2026)
CT-Eval: Benchmarking Chinese Text-to-Table Performance in Large Language Models
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
VCEval: Rethinking What is a Good Educational Video and How to Automatically Evaluate It
von: Zhu, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaoxuan, et al.
Veröffentlicht: (2024)
IDEATOR: Jailbreaking and Benchmarking Large Vision-Language Models Using Themselves
von: Wang, Ruofan, et al.
Veröffentlicht: (2024)
von: Wang, Ruofan, et al.
Veröffentlicht: (2024)
FiSMiness: A Finite State Machine Based Paradigm for Emotional Support Conversations
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
Can Pre-trained Language Models Understand Chinese Humor?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
XMeCap: Meme Caption Generation with Sub-Image Adaptability
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Fake Alignment: Are LLMs Really Aligned Well?
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
von: Wang, Yixu, et al.
Veröffentlicht: (2023)
Affective Flow Language Model for Emotional Support Conversation
von: Zou, Chenghui, et al.
Veröffentlicht: (2026)
von: Zou, Chenghui, et al.
Veröffentlicht: (2026)
Mitigating Unhelpfulness in Emotional Support Conversations with Multifaceted AI Feedback
von: Wang, Jiashuo, et al.
Veröffentlicht: (2024)
von: Wang, Jiashuo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
von: Li, Lingyu, et al.
Veröffentlicht: (2024) -
OVEL: Large Language Model as Memory Manager for Online Video Entity Linking
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024) -
HoneypotNet: Backdoor Attacks Against Model Extraction
von: Wang, Yixu, et al.
Veröffentlicht: (2025) -
MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models
von: Gu, Tianle, et al.
Veröffentlicht: (2024) -
The Other Mind: How Language Models Exhibit Human Temporal Cognition
von: Li, Lingyu, et al.
Veröffentlicht: (2025)