Minority-Aware Satisfaction Estimation in Dialogue Systems via Preference-Adaptive Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Fu, Yahui, Pang, Zi Haur, Kawahara, Tatsuya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Acknowledgment of Emotional States: Generating Validating Responses for Empathetic Dialogue
by: Pang, Zi Haur, et al.
Published: (2024)
by: Pang, Zi Haur, et al.
Published: (2024)
Paralinguistic Emotion-Aware Validation Timing Detection in Japanese Empathetic Spoken Dialogue
by: Pang, Zi Haur, et al.
Published: (2026)
by: Pang, Zi Haur, et al.
Published: (2026)
Human-Like Embodied AI Interviewer: Employing Android ERICA in Real International Conference
by: Pang, Zi Haur, et al.
Published: (2024)
by: Pang, Zi Haur, et al.
Published: (2024)
Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions?
by: Pang, Zi Haur, et al.
Published: (2025)
by: Pang, Zi Haur, et al.
Published: (2025)
StyEmp: Stylizing Empathetic Response Generation via Multi-Grained Prefix Encoder and Personality Reinforcement
by: Fu, Yahui, et al.
Published: (2024)
by: Fu, Yahui, et al.
Published: (2024)
Prompt-Guided Turn-Taking Prediction
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Enhancing Personality Recognition in Dialogue by Data Augmentation and Heterogeneous Conversational Graph Networks
by: Fu, Yahui, et al.
Published: (2024)
by: Fu, Yahui, et al.
Published: (2024)
Multilingual and Continuous Backchannel Prediction: A Cross-lingual Study
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Triadic Multi-party Voice Activity Projection for Turn-taking in Spoken Dialogue Systems
by: Elmers, Mikey, et al.
Published: (2025)
by: Elmers, Mikey, et al.
Published: (2025)
An Analysis of User Behaviors for Objectively Evaluating Spoken Dialogue Systems
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
CAUSE: Counterfactual Assessment of User Satisfaction Estimation in Task-Oriented Dialogue Systems
by: Abolghasemi, Amin, et al.
Published: (2024)
by: Abolghasemi, Amin, et al.
Published: (2024)
An LLM Benchmark for Addressee Recognition in Multi-modal Multi-party Dialogue
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
by: Shi, Hao, et al.
Published: (2024)
by: Shi, Hao, et al.
Published: (2024)
ERM-MinMaxGAP: Benchmarking and Mitigating Gender Bias in Multilingual Multimodal Speech-LLM Emotion Recognition
by: Pang, Zi Haur, et al.
Published: (2026)
by: Pang, Zi Haur, et al.
Published: (2026)
Tokenization Preference for Human and Machine Learning Model: An Annotation Study
by: Hiraoka, Tatsuya, et al.
Published: (2023)
by: Hiraoka, Tatsuya, et al.
Published: (2023)
LLM-guided Plan and Retrieval: A Strategic Alignment for Interpretable User Satisfaction Estimation in Dialogue
by: Kim, Sangyeop, et al.
Published: (2025)
by: Kim, Sangyeop, et al.
Published: (2025)
Analysis and Detection of Differences in Spoken User Behaviors between Autonomous and Wizard-of-Oz Systems
by: Elmers, Mikey, et al.
Published: (2024)
by: Elmers, Mikey, et al.
Published: (2024)
Efficient Preference-based Reinforcement Learning via Aligned Experience Estimation
by: Bai, Fengshuo, et al.
Published: (2024)
by: Bai, Fengshuo, et al.
Published: (2024)
Preference Distillation via Value based Reinforcement Learning
by: Kwon, Minchan, et al.
Published: (2025)
by: Kwon, Minchan, et al.
Published: (2025)
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Evaluating Task-Oriented Dialogue Consistency through Constraint Satisfaction
by: Labruna, Tiziano, et al.
Published: (2024)
by: Labruna, Tiziano, et al.
Published: (2024)
Why Do We Laugh? Annotation and Taxonomy Generation for Laughable Contexts in Spontaneous Text Conversation
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
by: Chaduvula, Sindhuja, et al.
Published: (2026)
by: Chaduvula, Sindhuja, et al.
Published: (2026)
Dialogue Systems for Emotional Support via Value Reinforcement
by: Kim, Juhee, et al.
Published: (2025)
by: Kim, Juhee, et al.
Published: (2025)
Evaluation of a semi-autonomous attentive listening system with takeover prompting
by: Kawai, Haruki, et al.
Published: (2024)
by: Kawai, Haruki, et al.
Published: (2024)
Distribution-Aware Reward Estimation for Test-Time Reinforcement Learning
by: Du, Bodong, et al.
Published: (2026)
by: Du, Bodong, et al.
Published: (2026)
Letting Tutor Personas "Speak Up" for LLMs: Learning Steering Vectors from Dialogue via Preference Optimization
by: Lee, Jaewook, et al.
Published: (2026)
by: Lee, Jaewook, et al.
Published: (2026)
Optimizing Conversational Quality in Spoken Dialogue Systems with Reinforcement Learning from AI Feedback
by: Arora, Siddhant, et al.
Published: (2026)
by: Arora, Siddhant, et al.
Published: (2026)
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
by: Chen, Jianan, et al.
Published: (2026)
by: Chen, Jianan, et al.
Published: (2026)
Interactive Dialogue Agents via Reinforcement Learning on Hindsight Regenerations
by: Hong, Joey, et al.
Published: (2024)
by: Hong, Joey, et al.
Published: (2024)
Enhancing Long-term RAG Chatbots with Psychological Models of Memory Importance and Forgetting
by: Sumida, Ryuichi, et al.
Published: (2024)
by: Sumida, Ryuichi, et al.
Published: (2024)
Adaptive Decoding via Latent Preference Optimization
by: Dhuliawala, Shehzaad, et al.
Published: (2024)
by: Dhuliawala, Shehzaad, et al.
Published: (2024)
Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
General Preference Reinforcement Learning
by: Umer, Muhammad, et al.
Published: (2026)
by: Umer, Muhammad, et al.
Published: (2026)
Reverse Engineering Human Preferences with Reinforcement Learning
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
Enhancing the Preference Extractor in Multi-turn Dialogues: From Annotating Disasters to Accurate Preference Extraction
by: Wang, Cheng, et al.
Published: (2025)
by: Wang, Cheng, et al.
Published: (2025)
Preference-Aware Rubric Learning for Personalized Evaluation
by: Qiu, Yilun, et al.
Published: (2026)
by: Qiu, Yilun, et al.
Published: (2026)
Aligning Large Language Model Behavior with Human Citation Preferences
by: Ando, Kenichiro, et al.
Published: (2026)
by: Ando, Kenichiro, et al.
Published: (2026)
Reinforcement Learning for Personalized Dialogue Management
by: Hengst, Floris den, et al.
Published: (2019)
by: Hengst, Floris den, et al.
Published: (2019)
Similar Items
-
Acknowledgment of Emotional States: Generating Validating Responses for Empathetic Dialogue
by: Pang, Zi Haur, et al.
Published: (2024) -
Paralinguistic Emotion-Aware Validation Timing Detection in Japanese Empathetic Spoken Dialogue
by: Pang, Zi Haur, et al.
Published: (2026) -
Human-Like Embodied AI Interviewer: Employing Android ERICA in Real International Conference
by: Pang, Zi Haur, et al.
Published: (2024) -
Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions?
by: Pang, Zi Haur, et al.
Published: (2025) -
StyEmp: Stylizing Empathetic Response Generation via Multi-Grained Prefix Encoder and Personality Reinforcement
by: Fu, Yahui, et al.
Published: (2024)