Accelerating Audio Research with Robotic Dummy Heads
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Austin, Sarkar, Kanad, Zhuang, Yongjie, Lin, Leo, Corey, Ryan M, Singer, Andrew C |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
by: Mishra, Ruchik, et al.
Published: (2024)
by: Mishra, Ruchik, et al.
Published: (2024)
Robotic Blended Sonification: Consequential Robot Sound as Creative Material for Human-Robot Interaction
by: Johansen, Stine S., et al.
Published: (2024)
by: Johansen, Stine S., et al.
Published: (2024)
Latent FxLMS: Accelerating Active Noise Control with Neural Adaptive Filters
by: Sarkar, Kanad, et al.
Published: (2025)
by: Sarkar, Kanad, et al.
Published: (2025)
Sound Judgment: Properties of Consequential Sounds Affecting Human-Perception of Robots
by: Allen, Aimee, et al.
Published: (2025)
by: Allen, Aimee, et al.
Published: (2025)
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
by: Hou, Yuanbo, et al.
Published: (2024)
by: Hou, Yuanbo, et al.
Published: (2024)
Examining Audio Communication Mechanisms for Supervising Fleets of Agricultural Robots
by: Kamboj, Abhi, et al.
Published: (2022)
by: Kamboj, Abhi, et al.
Published: (2022)
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
by: Li, Yuanchao, et al.
Published: (2023)
by: Li, Yuanchao, et al.
Published: (2023)
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025)
by: Tuttösí, Paige
Published: (2025)
Enhancing the NAO: Extending Capabilities of Legacy Robots for Long-Term Research
by: Wilson, Austin, et al.
Published: (2025)
by: Wilson, Austin, et al.
Published: (2025)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
by: Blum'e, Ashlae
Published: (2025)
by: Blum'e, Ashlae
Published: (2025)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
by: Wang, Hongbin, et al.
Published: (2025)
by: Wang, Hongbin, et al.
Published: (2025)
Single-Channel Robot Ego-Speech Filtering during Human-Robot Interaction
by: Li, Yue, et al.
Published: (2024)
by: Li, Yue, et al.
Published: (2024)
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
by: Chhaglani, Bhawana, et al.
Published: (2025)
by: Chhaglani, Bhawana, et al.
Published: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
by: Zheng, Shuoyang, et al.
Published: (2024)
by: Zheng, Shuoyang, et al.
Published: (2024)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
by: Li, Yinan, et al.
Published: (2026)
by: Li, Yinan, et al.
Published: (2026)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
by: Liu, Ailin, et al.
Published: (2024)
by: Liu, Ailin, et al.
Published: (2024)
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
by: Choi, Youngwon, et al.
Published: (2025)
by: Choi, Youngwon, et al.
Published: (2025)
Collaboration Between Robots, Interfaces and Humans: Practice-Based and Audience Perspectives
by: Savery, Anna, et al.
Published: (2024)
by: Savery, Anna, et al.
Published: (2024)
Optimizing Dysarthria Wake-Up Word Spotting: An End-to-End Approach for SLT 2024 LRDWWS Challenge
by: Liu, Shuiyun, et al.
Published: (2024)
by: Liu, Shuiyun, et al.
Published: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
by: Li, Yue, et al.
Published: (2024)
by: Li, Yue, et al.
Published: (2024)
MR-DAW: Towards Collaborative Digital Audio Workstations in Mixed Reality
by: Hopkins, Torin, et al.
Published: (2026)
by: Hopkins, Torin, et al.
Published: (2026)
ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection
by: Fan, Cunhang, et al.
Published: (2025)
by: Fan, Cunhang, et al.
Published: (2025)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
by: Li, Lu, et al.
Published: (2025)
by: Li, Lu, et al.
Published: (2025)
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
by: Zhou, Dongliang, et al.
Published: (2025)
by: Zhou, Dongliang, et al.
Published: (2025)
Directional Source Separation for Robust Speech Recognition on Smart Glasses
by: Feng, Tiantian, et al.
Published: (2023)
by: Feng, Tiantian, et al.
Published: (2023)
Human Perception of Audio Deepfakes
by: Müller, Nicolas M., et al.
Published: (2021)
by: Müller, Nicolas M., et al.
Published: (2021)
Head, posture, and full-body gestures in unscripted dyadic conversations in noise
by: Hládek, Ľuboš, et al.
Published: (2025)
by: Hládek, Ľuboš, et al.
Published: (2025)
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
by: Damacharla, Praveen, et al.
Published: (2023)
by: Damacharla, Praveen, et al.
Published: (2023)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
by: Piao, Ziyue, et al.
Published: (2024)
by: Piao, Ziyue, et al.
Published: (2024)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
by: Park, Seohyun, et al.
Published: (2025)
by: Park, Seohyun, et al.
Published: (2025)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
by: Han, Hyewon, et al.
Published: (2024)
by: Han, Hyewon, et al.
Published: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
by: Zhao, Yichun, et al.
Published: (2024)
by: Zhao, Yichun, et al.
Published: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
by: Yu, Luca Jiang-Tao, et al.
Published: (2024)
by: Yu, Luca Jiang-Tao, et al.
Published: (2024)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
by: Manukalpa, J. M. Chan Sri, et al.
Published: (2025)
by: Manukalpa, J. M. Chan Sri, et al.
Published: (2025)
Interfacing with history: Curating with audio augmented objects
by: Cliffe, Laurence
Published: (2024)
by: Cliffe, Laurence
Published: (2024)
Transhuman Ansambl - Voice Beyond Language
by: Ivsic, Lucija, et al.
Published: (2024)
by: Ivsic, Lucija, et al.
Published: (2024)
Cervical Auscultation Machine Learning for Dysphagia Assessment
by: Chia, An An, et al.
Published: (2024)
by: Chia, An An, et al.
Published: (2024)
ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds
by: Kobayashi, Atsuya, et al.
Published: (2020)
by: Kobayashi, Atsuya, et al.
Published: (2020)
Similar Items
-
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
by: Mishra, Ruchik, et al.
Published: (2024) -
Robotic Blended Sonification: Consequential Robot Sound as Creative Material for Human-Robot Interaction
by: Johansen, Stine S., et al.
Published: (2024) -
Latent FxLMS: Accelerating Active Noise Control with Neural Adaptive Filters
by: Sarkar, Kanad, et al.
Published: (2025) -
Sound Judgment: Properties of Consequential Sounds Affecting Human-Perception of Robots
by: Allen, Aimee, et al.
Published: (2025) -
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
by: Hou, Yuanbo, et al.
Published: (2024)