Active Listener: Continuous Generation of Listener's Head Motion Response in Dyadic Interactions
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ghosh, Bishal, Li, Emma, Guha, Tanaya |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CustomListener: Text-guided Responsive Interaction for User-friendly Listening Head Generation
par: Liu, Xi, et autres
Publié: (2024)
par: Liu, Xi, et autres
Publié: (2024)
Listen, Think, and Understand
par: Gong, Yuan, et autres
Publié: (2023)
par: Gong, Yuan, et autres
Publié: (2023)
I Know You're Listening: Adaptive Voice for HRI
par: Tuttösí, Paige
Publié: (2025)
par: Tuttösí, Paige
Publié: (2025)
RF-GML: Reference-Free Generative Machine Listener
par: Biswas, Arijit, et autres
Publié: (2024)
par: Biswas, Arijit, et autres
Publié: (2024)
A Multi-loudspeaker Binaural Room Impulse Response Dataset with High-Resolution Translational and Rotational Head Coordinates in a Listening Room
par: Qiao, Yue, et autres
Publié: (2024)
par: Qiao, Yue, et autres
Publié: (2024)
Requirements for Mass Adoption of Assistive Listening Technology by the General Public
par: Kaufmann, Thomas B., et autres
Publié: (2023)
par: Kaufmann, Thomas B., et autres
Publié: (2023)
DIFFA: Large Language Diffusion Models Can Listen and Understand
par: Zhou, Jiaming, et autres
Publié: (2025)
par: Zhou, Jiaming, et autres
Publié: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
par: Chung, Soo-Whan, et autres
Publié: (2025)
par: Chung, Soo-Whan, et autres
Publié: (2025)
Evaluating Speech Enhancement Systems Through Listening Effort
par: Gelderblom, Femke B., et autres
Publié: (2024)
par: Gelderblom, Femke B., et autres
Publié: (2024)
Listen First, Then Answer: Timestamp-Grounded Speech Reasoning
par: Jeong, Jihoon, et autres
Publié: (2026)
par: Jeong, Jihoon, et autres
Publié: (2026)
Listen to Extract: Onset-Prompted Target Speaker Extraction
par: Shen, Pengjie, et autres
Publié: (2025)
par: Shen, Pengjie, et autres
Publié: (2025)
Listening to Multi-talker Conversations: Modular and End-to-end Perspectives
par: Raj, Desh
Publié: (2024)
par: Raj, Desh
Publié: (2024)
Reproducing the Acoustic Velocity Vectors in a Circular Listening Area
par: Wang, Jiarui, et autres
Publié: (2024)
par: Wang, Jiarui, et autres
Publié: (2024)
Listening broadband physical model for microphones: a first step
par: Millot, Laurent, et autres
Publié: (2024)
par: Millot, Laurent, et autres
Publié: (2024)
Joint Minimum Processing Beamforming and Near-end Listening Enhancement
par: Fuglsig, Andreas J., et autres
Publié: (2023)
par: Fuglsig, Andreas J., et autres
Publié: (2023)
Unifying Listener Scoring Scales: Comparison Learning Framework for Speech Quality Assessment and Continuous Speech Emotion Recognition
par: Hu, Cheng-Hung, et autres
Publié: (2025)
par: Hu, Cheng-Hung, et autres
Publié: (2025)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
par: Chen, Zhengyang, et autres
Publié: (2024)
par: Chen, Zhengyang, et autres
Publié: (2024)
Listen and Move: Improving GANs Coherency in Agnostic Sound-to-Video Generation
par: Redondo, Rafael
Publié: (2024)
par: Redondo, Rafael
Publié: (2024)
What Do Neurons Listen To? A Neuron-level Dissection of a General-purpose Audio Model
par: Kawamura, Takao, et autres
Publié: (2026)
par: Kawamura, Takao, et autres
Publié: (2026)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
par: Shen, Yu-Han, et autres
Publié: (2018)
par: Shen, Yu-Han, et autres
Publié: (2018)
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
par: Yamamoto, Katsuhiko, et autres
Publié: (2025)
par: Yamamoto, Katsuhiko, et autres
Publié: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
par: Rahimi, Akam, et autres
Publié: (2025)
par: Rahimi, Akam, et autres
Publié: (2025)
Listenable Maps for Audio Classifiers
par: Paissan, Francesco, et autres
Publié: (2024)
par: Paissan, Francesco, et autres
Publié: (2024)
Listening Between the Lines: Synthetic Speech Detection Disregarding Verbal Content
par: Salvi, Davide, et autres
Publié: (2024)
par: Salvi, Davide, et autres
Publié: (2024)
Long-Term, Store-Front Robotics: Interactive Music for Robotic Arm, Caxixi and Frame Drums
par: Savery, Richard, et autres
Publié: (2024)
par: Savery, Richard, et autres
Publié: (2024)
Can Masked Autoencoders Also Listen to Birds?
par: Rauch, Lukas, et autres
Publié: (2025)
par: Rauch, Lukas, et autres
Publié: (2025)
OCR-Enhanced Multimodal ASR Can Read While Listening
par: Chen, Junli, et autres
Publié: (2026)
par: Chen, Junli, et autres
Publié: (2026)
Deep CLAS: Deep Contextual Listen, Attend and Spell
par: Wang, Mengzhi, et autres
Publié: (2024)
par: Wang, Mengzhi, et autres
Publié: (2024)
Development of the Listening in Spatialized Noise-Sentences (LiSN-S) Test in Brazilian Portuguese: Presentation Software, Speech Stimuli, and Sentence Equivalence
par: Masiero, Bruno S., et autres
Publié: (2024)
par: Masiero, Bruno S., et autres
Publié: (2024)
Spatial Analysis and Synthesis Methods: Subjective and Objective Evaluations Using Various Microphone Arrays in the Auralization of a Critical Listening Room
par: Pawlak, Alan, et autres
Publié: (2024)
par: Pawlak, Alan, et autres
Publié: (2024)
Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
par: Xiao, Yang, et autres
Publié: (2025)
par: Xiao, Yang, et autres
Publié: (2025)
A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses
par: Maghsoudi, Maryam, et autres
Publié: (2025)
par: Maghsoudi, Maryam, et autres
Publié: (2025)
Listening and Seeing Again: Generative Error Correction for Audio-Visual Speech Recognition
par: Liu, Rui, et autres
Publié: (2025)
par: Liu, Rui, et autres
Publié: (2025)
Listenable Maps for Zero-Shot Audio Classifiers
par: Paissan, Francesco, et autres
Publié: (2024)
par: Paissan, Francesco, et autres
Publié: (2024)
Listen, Chat, and Remix: Text-Guided Soundscape Remixing for Enhanced Auditory Experience
par: Jiang, Xilin, et autres
Publié: (2024)
par: Jiang, Xilin, et autres
Publié: (2024)
A Data-Centric Framework for Machine Listening Projects: Addressing Large-Scale Data Acquisition and Labeling through Active Learning
par: Naranjo-Alcazar, Javier, et autres
Publié: (2024)
par: Naranjo-Alcazar, Javier, et autres
Publié: (2024)
Leveraging Multiple Speech Enhancers for Non-Intrusive Intelligibility Prediction for Hearing-Impaired Listeners
par: Cao, Boxuan, et autres
Publié: (2025)
par: Cao, Boxuan, et autres
Publié: (2025)
Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System
par: Kato, Kazushi, et autres
Publié: (2025)
par: Kato, Kazushi, et autres
Publié: (2025)
Theoretical Framework for the Optimization of Microphone Array Configuration for Humanoid Robot Audition
par: Tourbabin, Vladimir, et autres
Publié: (2024)
par: Tourbabin, Vladimir, et autres
Publié: (2024)
Disentangled Acoustic Fields For Multimodal Physical Scene Understanding
par: Yin, Jie, et autres
Publié: (2024)
par: Yin, Jie, et autres
Publié: (2024)
Documents similaires
-
CustomListener: Text-guided Responsive Interaction for User-friendly Listening Head Generation
par: Liu, Xi, et autres
Publié: (2024) -
Listen, Think, and Understand
par: Gong, Yuan, et autres
Publié: (2023) -
I Know You're Listening: Adaptive Voice for HRI
par: Tuttösí, Paige
Publié: (2025) -
RF-GML: Reference-Free Generative Machine Listener
par: Biswas, Arijit, et autres
Publié: (2024) -
A Multi-loudspeaker Binaural Room Impulse Response Dataset with High-Resolution Translational and Rotational Head Coordinates in a Listening Room
par: Qiao, Yue, et autres
Publié: (2024)