Paralinguistic Emotion-Aware Validation Timing Detection in Japanese Empathetic Spoken Dialogue
Fuente:
arXiv
Guardado en:
| Autores principales: | Pang, Zi Haur, Fu, Yahui, Gao, Yuan, Kawahara, Tatsuya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Acknowledgment of Emotional States: Generating Validating Responses for Empathetic Dialogue
por: Pang, Zi Haur, et al.
Publicado: (2024)
por: Pang, Zi Haur, et al.
Publicado: (2024)
ERM-MinMaxGAP: Benchmarking and Mitigating Gender Bias in Multilingual Multimodal Speech-LLM Emotion Recognition
por: Pang, Zi Haur, et al.
Publicado: (2026)
por: Pang, Zi Haur, et al.
Publicado: (2026)
Minority-Aware Satisfaction Estimation in Dialogue Systems via Preference-Adaptive Reinforcement Learning
por: Fu, Yahui, et al.
Publicado: (2025)
por: Fu, Yahui, et al.
Publicado: (2025)
Prompt-Guided Turn-Taking Prediction
por: Inoue, Koji, et al.
Publicado: (2025)
por: Inoue, Koji, et al.
Publicado: (2025)
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
por: Gao, Yuan, et al.
Publicado: (2025)
por: Gao, Yuan, et al.
Publicado: (2025)
Multilingual and Continuous Backchannel Prediction: A Cross-lingual Study
por: Inoue, Koji, et al.
Publicado: (2025)
por: Inoue, Koji, et al.
Publicado: (2025)
OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue
por: Geng, Xuelong, et al.
Publicado: (2025)
por: Geng, Xuelong, et al.
Publicado: (2025)
Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions?
por: Pang, Zi Haur, et al.
Publicado: (2025)
por: Pang, Zi Haur, et al.
Publicado: (2025)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
por: Nakata, Wataru, et al.
Publicado: (2024)
por: Nakata, Wataru, et al.
Publicado: (2024)
Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data
por: Xie, Jingran, et al.
Publicado: (2025)
por: Xie, Jingran, et al.
Publicado: (2025)
GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness
por: Chen, Hongjie, et al.
Publicado: (2025)
por: Chen, Hongjie, et al.
Publicado: (2025)
Two-Stage Acoustic Adaptation with Gated Cross-Attention Adapters for LLM-Based Multi-Talker Speech Recognition
por: Shi, Hao, et al.
Publicado: (2026)
por: Shi, Hao, et al.
Publicado: (2026)
Semantic-Aware Interruption Detection in Spoken Dialogue Systems: Benchmark, Metric, and Model
por: Xia, Kangxiang, et al.
Publicado: (2026)
por: Xia, Kangxiang, et al.
Publicado: (2026)
DeepDialogue: A Multi-Turn Emotionally-Rich Spoken Dialogue Dataset
por: Koudounas, Alkis, et al.
Publicado: (2025)
por: Koudounas, Alkis, et al.
Publicado: (2025)
E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models
por: Xue, Hongfei, et al.
Publicado: (2023)
por: Xue, Hongfei, et al.
Publicado: (2023)
Are Paralinguistic Representations all that is needed for Speech Emotion Recognition?
por: Phukan, Orchid Chetia, et al.
Publicado: (2024)
por: Phukan, Orchid Chetia, et al.
Publicado: (2024)
StyEmp: Stylizing Empathetic Response Generation via Multi-Grained Prefix Encoder and Personality Reinforcement
por: Fu, Yahui, et al.
Publicado: (2024)
por: Fu, Yahui, et al.
Publicado: (2024)
Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox
por: Pang, Jiacheng, et al.
Publicado: (2026)
por: Pang, Jiacheng, et al.
Publicado: (2026)
VoxMind: An End-to-End Agentic Spoken Dialogue System
por: Liang, Tianle, et al.
Publicado: (2026)
por: Liang, Tianle, et al.
Publicado: (2026)
Serialized Speech Information Guidance with Overlapped Encoding Separation for Multi-Speaker Automatic Speech Recognition
por: Shi, Hao, et al.
Publicado: (2024)
por: Shi, Hao, et al.
Publicado: (2024)
SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation
por: Liu, Ruohan, et al.
Publicado: (2026)
por: Liu, Ruohan, et al.
Publicado: (2026)
MOSS-TTSD: Text to Spoken Dialogue Generation
por: Zhang, Yuqian, et al.
Publicado: (2026)
por: Zhang, Yuqian, et al.
Publicado: (2026)
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
por: Inoue, Koji, et al.
Publicado: (2025)
por: Inoue, Koji, et al.
Publicado: (2025)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
por: Shi, Hao, et al.
Publicado: (2024)
por: Shi, Hao, et al.
Publicado: (2024)
Benchmarking Japanese Speech Recognition on ASR-LLM Setups with Multi-Pass Augmented Generative Error Correction
por: Ko, Yuka, et al.
Publicado: (2024)
por: Ko, Yuka, et al.
Publicado: (2024)
An LLM Benchmark for Addressee Recognition in Multi-modal Multi-party Dialogue
por: Inoue, Koji, et al.
Publicado: (2025)
por: Inoue, Koji, et al.
Publicado: (2025)
SONAR: Self-Distilled Continual Pre-training for Domain Adaptive Audio Representation
por: Zhang, Yizhou, et al.
Publicado: (2025)
por: Zhang, Yizhou, et al.
Publicado: (2025)
Human-Like Embodied AI Interviewer: Employing Android ERICA in Real International Conference
por: Pang, Zi Haur, et al.
Publicado: (2024)
por: Pang, Zi Haur, et al.
Publicado: (2024)
Resurfacing Paralinguistic Awareness in Large Audio Language Models
por: Yang, Hao, et al.
Publicado: (2026)
por: Yang, Hao, et al.
Publicado: (2026)
SingingSDS: A Singing-Capable Spoken Dialogue System for Conversational Roleplay Applications
por: Han, Jionghao, et al.
Publicado: (2025)
por: Han, Jionghao, et al.
Publicado: (2025)
Efficient and Robust Long-Form Speech Recognition with Hybrid H3-Conformer
por: Honda, Tomoki, et al.
Publicado: (2024)
por: Honda, Tomoki, et al.
Publicado: (2024)
Reflecting Twice before Speaking with Empathy: Self-Reflective Alternating Inference for Empathy-Aware End-to-End Spoken Dialogue
por: Jia, Yuhang, et al.
Publicado: (2026)
por: Jia, Yuhang, et al.
Publicado: (2026)
Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
por: Kim, Heeseung, et al.
Publicado: (2024)
por: Kim, Heeseung, et al.
Publicado: (2024)
Towards Machine Unlearning for Paralinguistic Speech Processing
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
por: Li, Yuanchao, et al.
Publicado: (2023)
por: Li, Yuanchao, et al.
Publicado: (2023)
Optimizing Conversational Quality in Spoken Dialogue Systems with Reinforcement Learning from AI Feedback
por: Arora, Siddhant, et al.
Publicado: (2026)
por: Arora, Siddhant, et al.
Publicado: (2026)
Geolocation-Aware Robust Spoken Language Identification
por: Wang, Qingzheng, et al.
Publicado: (2025)
por: Wang, Qingzheng, et al.
Publicado: (2025)
FlashLabs Chroma 1.0: A Real-Time End-to-End Spoken Dialogue Model with Personalized Voice Cloning
por: Chen, Tanyu, et al.
Publicado: (2026)
por: Chen, Tanyu, et al.
Publicado: (2026)
RE-LLM: Refining Empathetic Speech-LLM Responses by Integrating Emotion Nuance
por: Chen, Jing-Han, et al.
Publicado: (2026)
por: Chen, Jing-Han, et al.
Publicado: (2026)
Combining Deterministic Enhanced Conditions with Dual-Streaming Encoding for Diffusion-Based Speech Enhancement
por: Shi, Hao, et al.
Publicado: (2025)
por: Shi, Hao, et al.
Publicado: (2025)
Ejemplares similares
-
Acknowledgment of Emotional States: Generating Validating Responses for Empathetic Dialogue
por: Pang, Zi Haur, et al.
Publicado: (2024) -
ERM-MinMaxGAP: Benchmarking and Mitigating Gender Bias in Multilingual Multimodal Speech-LLM Emotion Recognition
por: Pang, Zi Haur, et al.
Publicado: (2026) -
Minority-Aware Satisfaction Estimation in Dialogue Systems via Preference-Adaptive Reinforcement Learning
por: Fu, Yahui, et al.
Publicado: (2025) -
Prompt-Guided Turn-Taking Prediction
por: Inoue, Koji, et al.
Publicado: (2025) -
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
por: Gao, Yuan, et al.
Publicado: (2025)