Talk With Human-like Agents: Empathetic Dialogue Through Perceptible Acoustic Reception and Reaction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yan, Haoqiu, Zhu, Yongxin, Zheng, Kai, Liu, Bing, Cao, Haoyu, Jiang, Deqiang, Xu, Linli |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data
par: Xie, Jingran, et autres
Publié: (2025)
par: Xie, Jingran, et autres
Publié: (2025)
Chain-Talker: Chain Understanding and Rendering for Empathetic Conversational Speech Synthesis
par: Hu, Yifan, et autres
Publié: (2025)
par: Hu, Yifan, et autres
Publié: (2025)
Generative Pre-trained Speech Language Model with Efficient Hierarchical Transformer
par: Zhu, Yongxin, et autres
Publié: (2024)
par: Zhu, Yongxin, et autres
Publié: (2024)
IMPACT: Industrial Machine Perception via Acoustic Cognitive Transformer
par: Han, Changheon, et autres
Publié: (2025)
par: Han, Changheon, et autres
Publié: (2025)
Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Multi-Channel Conformer
par: Yang, Bing, et autres
Publié: (2023)
par: Yang, Bing, et autres
Publié: (2023)
Acoustic Volume Rendering for Neural Impulse Response Fields
par: Lan, Zitong, et autres
Publié: (2024)
par: Lan, Zitong, et autres
Publié: (2024)
MFA-KWS: Effective Keyword Spotting with Multi-head Frame-asynchronous Decoding
par: Xi, Yu, et autres
Publié: (2025)
par: Xi, Yu, et autres
Publié: (2025)
Acoustic BPE for Speech Generation with Discrete Tokens
par: Shen, Feiyu, et autres
Publié: (2023)
par: Shen, Feiyu, et autres
Publié: (2023)
Data-Efficient Low-Complexity Acoustic Scene Classification via Distilling and Progressive Pruning
par: Han, Bing, et autres
Publié: (2024)
par: Han, Bing, et autres
Publié: (2024)
DialogueSidon: Recovering Full-Duplex Dialogue Tracks from In-the-Wild Dialogue Audio
par: Nakata, Wataru, et autres
Publié: (2026)
par: Nakata, Wataru, et autres
Publié: (2026)
Comprehend and Talk: Text to Speech Synthesis via Dual Language Modeling
par: Cao, Junjie, et autres
Publié: (2025)
par: Cao, Junjie, et autres
Publié: (2025)
Masked Self-distilled Transducer-based Keyword Spotting with Semi-autoregressive Decoding
par: Xi, Yu, et autres
Publié: (2025)
par: Xi, Yu, et autres
Publié: (2025)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature
par: Du, Chenpeng, et autres
Publié: (2022)
par: Du, Chenpeng, et autres
Publié: (2022)
Evaluation of Virtual Acoustic Environments with Different Acoustic Level of Detail
par: Fichna, Stefan, et autres
Publié: (2023)
par: Fichna, Stefan, et autres
Publié: (2023)
Polyphonia: Zero-Shot Timbre Transfer in Polyphonic Music with Acoustic-Informed Attention Calibration
par: Li, Haowen, et autres
Publié: (2026)
par: Li, Haowen, et autres
Publié: (2026)
Cross-Talk Speech Reduction, by Separation, for Separation
par: Wang, Zhong-Qiu, et autres
Publié: (2026)
par: Wang, Zhong-Qiu, et autres
Publié: (2026)
From Human Speech to Ocean Signals: Transferring Speech Large Models for Underwater Acoustic Target Recognition
par: Huang, Mengcheng, et autres
Publié: (2026)
par: Huang, Mengcheng, et autres
Publié: (2026)
BLSP-Emo: Towards Empathetic Large Speech-Language Models
par: Wang, Chen, et autres
Publié: (2024)
par: Wang, Chen, et autres
Publié: (2024)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
par: Tailleur, Modan, et autres
Publié: (2024)
par: Tailleur, Modan, et autres
Publié: (2024)
Exploring Differences between Human Perception and Model Inference in Audio Event Recognition
par: Tan, Yizhou, et autres
Publié: (2024)
par: Tan, Yizhou, et autres
Publié: (2024)
Pianoroll-Event: A Novel Score Representation for Symbolic Music
par: Qian, Lekai, et autres
Publié: (2026)
par: Qian, Lekai, et autres
Publié: (2026)
Enroll-on-Wakeup: A First Comparative Study of Target Speech Extraction for Seamless Interaction in Real Noisy Human-Machine Dialogue Scenarios
par: Yang, Yiming, et autres
Publié: (2026)
par: Yang, Yiming, et autres
Publié: (2026)
SAC: Neural Speech Codec with Semantic-Acoustic Dual-Stream Quantization
par: Chen, Wenxi, et autres
Publié: (2025)
par: Chen, Wenxi, et autres
Publié: (2025)
Towards General Discrete Speech Codec for Complex Acoustic Environments: A Study of Reconstruction and Downstream Task Consistency
par: Wang, Haoran, et autres
Publié: (2025)
par: Wang, Haoran, et autres
Publié: (2025)
Sound Field Synthesis with Acoustic Waves
par: Mansour, Mohamed F.
Publié: (2024)
par: Mansour, Mohamed F.
Publié: (2024)
Continual Learning for Acoustic Event Classification
par: Xiao, Yang
Publié: (2025)
par: Xiao, Yang
Publié: (2025)
SonicSense: Object Perception from In-Hand Acoustic Vibration
par: Liu, Jiaxun, et autres
Publié: (2024)
par: Liu, Jiaxun, et autres
Publié: (2024)
RE-LLM: Refining Empathetic Speech-LLM Responses by Integrating Emotion Nuance
par: Chen, Jing-Han, et autres
Publié: (2026)
par: Chen, Jing-Han, et autres
Publié: (2026)
ParaLBench: A Large-Scale Benchmark for Computational Paralinguistics over Acoustic Foundation Models
par: Zhang, Zixing, et autres
Publié: (2024)
par: Zhang, Zixing, et autres
Publié: (2024)
TDT-KWS: Fast And Accurate Keyword Spotting Using Token-and-duration Transducer
par: Xi, Yu, et autres
Publié: (2024)
par: Xi, Yu, et autres
Publié: (2024)
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
par: Wen, Wen, et autres
Publié: (2024)
par: Wen, Wen, et autres
Publié: (2024)
Semantic Proximity Alignment: Towards Human Perception-consistent Audio Tagging by Aligning with Label Text Description
par: Liu, Wuyang, et autres
Publié: (2023)
par: Liu, Wuyang, et autres
Publié: (2023)
XANE: eXplainable Acoustic Neural Embeddings
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
Neural Kalman Filters for Acoustic Echo Cancellation
par: Seidel, Ernst, et autres
Publié: (2025)
par: Seidel, Ernst, et autres
Publié: (2025)
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
par: Hussein, Amir, et autres
Publié: (2025)
par: Hussein, Amir, et autres
Publié: (2025)
FastTurn: Unifying Acoustic and Streaming Semantic Cues for Low-Latency and Robust Turn Detection
par: Wang, Chengyou, et autres
Publié: (2026)
par: Wang, Chengyou, et autres
Publié: (2026)
Theoretical Model of Acoustic Power Transfer Through Solids
par: Kochliaridis, Ippokratis, et autres
Publié: (2025)
par: Kochliaridis, Ippokratis, et autres
Publié: (2025)
On Time Delay Interpolation for Improved Acoustic Reflector Localization
par: Rosseel, Hannes, et autres
Publié: (2025)
par: Rosseel, Hannes, et autres
Publié: (2025)
Production and Manufacturing of 3D Printed Acoustic Guitars
par: Tran, Timothy, et autres
Publié: (2025)
par: Tran, Timothy, et autres
Publié: (2025)
Documents similaires
-
Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data
par: Xie, Jingran, et autres
Publié: (2025) -
Chain-Talker: Chain Understanding and Rendering for Empathetic Conversational Speech Synthesis
par: Hu, Yifan, et autres
Publié: (2025) -
Generative Pre-trained Speech Language Model with Efficient Hierarchical Transformer
par: Zhu, Yongxin, et autres
Publié: (2024) -
IMPACT: Industrial Machine Perception via Acoustic Cognitive Transformer
par: Han, Changheon, et autres
Publié: (2025) -
Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Multi-Channel Conformer
par: Yang, Bing, et autres
Publié: (2023)