DialSim: A Dialogue Simulator for Evaluating Long-Term Multi-Party Dialogue Understanding of Conversational Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Jiho, Chay, Woosog, Hwang, Hyeonji, Kyung, Daeun, Chung, Hyunseung, Cho, Eunbyeol, Kwon, Yeonsu, Jo, Yohan, Choi, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant Simulation
by: Kim, Jiho, et al.
Published: (2025)
by: Kim, Jiho, et al.
Published: (2025)
ECG-Agent: On-Device Tool-Calling Agent for ECG Multi-Turn Dialogue
by: Chung, Hyunseung, et al.
Published: (2026)
by: Chung, Hyunseung, et al.
Published: (2026)
Time is Not Enough: Time-Frequency based Explanation for Time-Series Black-Box Models
by: Chung, Hyunseung, et al.
Published: (2024)
by: Chung, Hyunseung, et al.
Published: (2024)
SCARE: A Benchmark for SQL Correction and Question Answerability Classification for Reliable EHR Question Answering
by: Lee, Gyubok, et al.
Published: (2025)
by: Lee, Gyubok, et al.
Published: (2025)
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
by: Lee, Gyubok, et al.
Published: (2024)
by: Lee, Gyubok, et al.
Published: (2024)
PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions
by: Kyung, Daeun, et al.
Published: (2025)
by: Kyung, Daeun, et al.
Published: (2025)
ToolDial: Multi-turn Dialogue Generation Method for Tool-Augmented Language Models
by: Shim, Jeonghoon, et al.
Published: (2025)
by: Shim, Jeonghoon, et al.
Published: (2025)
EHRCon: Dataset for Checking Consistency between Unstructured Notes and Structured Tables in Electronic Health Records
by: Kwon, Yeonsu, et al.
Published: (2024)
by: Kwon, Yeonsu, et al.
Published: (2024)
KMI: A Dataset of Korean Motivational Interviewing Dialogues for Psychotherapy
by: Kim, Hyunjong, et al.
Published: (2025)
by: Kim, Hyunjong, et al.
Published: (2025)
From Conversation to Query Execution: Benchmarking User and Tool Interactions for EHR Database Agents
by: Lee, Gyubok, et al.
Published: (2025)
by: Lee, Gyubok, et al.
Published: (2025)
Improving Dialogue State Tracking through Combinatorial Search for In-Context Examples
by: Pyun, Haesung, et al.
Published: (2025)
by: Pyun, Haesung, et al.
Published: (2025)
CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment
by: Shayanfar, Radin, et al.
Published: (2025)
by: Shayanfar, Radin, et al.
Published: (2025)
Towards Lifelong Dialogue Agents via Timeline-based Memory Management
by: Ong, Kai Tzu-iunn, et al.
Published: (2024)
by: Ong, Kai Tzu-iunn, et al.
Published: (2024)
EHRSQL: A Practical Text-to-SQL Benchmark for Electronic Health Records
by: Lee, Gyubok, et al.
Published: (2023)
by: Lee, Gyubok, et al.
Published: (2023)
SpokenUS: A Spoken User Simulator for Task-Oriented Dialogue
by: Lee, Jonggeun, et al.
Published: (2026)
by: Lee, Jonggeun, et al.
Published: (2026)
ComperDial: Commonsense Persona-grounded Dialogue Dataset and Benchmark
by: Wakaki, Hiromi, et al.
Published: (2024)
by: Wakaki, Hiromi, et al.
Published: (2024)
Dialogue Systems for Emotional Support via Value Reinforcement
by: Kim, Juhee, et al.
Published: (2025)
by: Kim, Juhee, et al.
Published: (2025)
Understanding Human Perception of Music Plagiarism Through a Computational Approach
by: Hwang, Daeun, et al.
Published: (2026)
by: Hwang, Daeun, et al.
Published: (2026)
EarthDial: Turning Multi-sensory Earth Observations to Interactive Dialogues
by: Soni, Sagar, et al.
Published: (2024)
by: Soni, Sagar, et al.
Published: (2024)
Generating Multi-Table Time Series EHR from Latent Space with Minimal Preprocessing
by: Cho, Eunbyeol, et al.
Published: (2025)
by: Cho, Eunbyeol, et al.
Published: (2025)
SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?
by: Lee, Jonggeun, et al.
Published: (2026)
by: Lee, Jonggeun, et al.
Published: (2026)
LEEETs-Dial: Linguistic Entrainment in End-to-End Task-oriented Dialogue systems
by: Kumar, Nalin, et al.
Published: (2023)
by: Kumar, Nalin, et al.
Published: (2023)
FineDialFact: A benchmark for Fine-grained Dialogue Fact Verification
by: Chen, Xiangyan, et al.
Published: (2025)
by: Chen, Xiangyan, et al.
Published: (2025)
Towards Predicting Temporal Changes in a Patient's Chest X-ray Images based on Electronic Health Records
by: Kyung, Daeun, et al.
Published: (2024)
by: Kyung, Daeun, et al.
Published: (2024)
DialogueAgents: A Hybrid Agent-Based Speech Synthesis Framework for Multi-Party Dialogue
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Pre-Storage Reasoning for Episodic Memory: Shifting Inference Burden to Memory for Personalized Dialogue
by: Kim, Sangyeop, et al.
Published: (2025)
by: Kim, Sangyeop, et al.
Published: (2025)
R2-KG: General-Purpose Dual-Agent Framework for Reliable Reasoning on Knowledge Graphs
by: Jo, Sumin, et al.
Published: (2025)
by: Jo, Sumin, et al.
Published: (2025)
ProKG-Dial: Progressive Multi-Turn Dialogue Construction with Domain Knowledge Graphs
by: Liang, Yuanyuan, et al.
Published: (2025)
by: Liang, Yuanyuan, et al.
Published: (2025)
Dial-MAE: ConTextual Masked Auto-Encoder for Retrieval-based Dialogue Systems
by: Su, Zhenpeng, et al.
Published: (2023)
by: Su, Zhenpeng, et al.
Published: (2023)
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
by: Luo, Wen, et al.
Published: (2024)
by: Luo, Wen, et al.
Published: (2024)
Dial-In LLM: Human-Aligned LLM-in-the-loop Intent Clustering for Customer Service Dialogues
by: Hong, Mengze, et al.
Published: (2024)
by: Hong, Mengze, et al.
Published: (2024)
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation
by: Wang, Xiaoyu, et al.
Published: (2024)
by: Wang, Xiaoyu, et al.
Published: (2024)
Dialogue for Intercultural Understanding
Published: (2021)
Published: (2021)
Adapting Text-based Dialogue State Tracker for Spoken Dialogues
by: Yoon, Jaeseok, et al.
Published: (2023)
by: Yoon, Jaeseok, et al.
Published: (2023)
MindDial: Belief Dynamics Tracking with Theory-of-Mind Modeling for Situated Neural Dialogue Generation
by: Qiu, Shuwen, et al.
Published: (2023)
by: Qiu, Shuwen, et al.
Published: (2023)
HealthDial: A No-Code LLM-Assisted Dialogue Authoring Tool for Healthcare Virtual Agents
by: Nouraei, Farnaz, et al.
Published: (2025)
by: Nouraei, Farnaz, et al.
Published: (2025)
DialBGM: A Benchmark for Background Music Recommendation from Everyday Multi-Turn Dialogues
by: Shin, Joonhyeok, et al.
Published: (2026)
by: Shin, Joonhyeok, et al.
Published: (2026)
EduDial: Constructing a Large-scale Multi-turn Teacher-Student Dialogue Corpus
by: Wei, Shouang, et al.
Published: (2025)
by: Wei, Shouang, et al.
Published: (2025)
Evaluating Long-Horizon Memory for Multi-Party Collaborative Dialogues
by: Hu, Chuanrui, et al.
Published: (2026)
by: Hu, Chuanrui, et al.
Published: (2026)
Better Not to Propagate: Understanding Edge Uncertainty and Over-smoothing in Signed Graph Neural Networks
by: Choi, Yoonhyuk, et al.
Published: (2024)
by: Choi, Yoonhyuk, et al.
Published: (2024)
Similar Items
-
ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant Simulation
by: Kim, Jiho, et al.
Published: (2025) -
ECG-Agent: On-Device Tool-Calling Agent for ECG Multi-Turn Dialogue
by: Chung, Hyunseung, et al.
Published: (2026) -
Time is Not Enough: Time-Frequency based Explanation for Time-Series Black-Box Models
by: Chung, Hyunseung, et al.
Published: (2024) -
SCARE: A Benchmark for SQL Correction and Question Answerability Classification for Reliable EHR Question Answering
by: Lee, Gyubok, et al.
Published: (2025) -
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
by: Lee, Gyubok, et al.
Published: (2024)