Uncovering the role of semantic and acoustic cues in normal and dichotic listening
Fuente:
arXiv
Saved in:
| Main Authors: | Kankanala, Sai Samrat, Soman, Akshara, Ganapathy, Sriram |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Humans and Machines on Complex Multilingual Speech Understanding Tasks
by: Kankanala, Sai Samrat, et al.
Published: (2025)
by: Kankanala, Sai Samrat, et al.
Published: (2025)
Summary of the DISPLACE Challenge 2023 - DIarization of SPeaker and LAnguage in Conversational Environments
by: Baghel, Shikha, et al.
Published: (2023)
by: Baghel, Shikha, et al.
Published: (2023)
On fusing active and passive acoustic sensing for simultaneous localization and mapping
by: Bradley, Aidan J., et al.
Published: (2024)
by: Bradley, Aidan J., et al.
Published: (2024)
improving partition-block-based acoustic echo canceler in under-modeling scenarios
by: Fan, Wenzhi, et al.
Published: (2020)
by: Fan, Wenzhi, et al.
Published: (2020)
RIR-Mega: a large-scale simulated room impulse response dataset for machine learning and room acoustics modeling
by: Goswami, Mandip
Published: (2025)
by: Goswami, Mandip
Published: (2025)
Human perception of audio deepfakes: the role of language and speaking style
by: Segundo, Eugenia San, et al.
Published: (2025)
by: Segundo, Eugenia San, et al.
Published: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
by: Dutta, Soumya, et al.
Published: (2024)
by: Dutta, Soumya, et al.
Published: (2024)
Deep learning classification system for coconut maturity levels based on acoustic signals
by: Caladcad, June Anne, et al.
Published: (2024)
by: Caladcad, June Anne, et al.
Published: (2024)
Passive acoustic non-line-of-sight localization without a relay surface
by: Sommer, Tal I., et al.
Published: (2025)
by: Sommer, Tal I., et al.
Published: (2025)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
by: Waheed, Muhammad Fasih, et al.
Published: (2026)
by: Waheed, Muhammad Fasih, et al.
Published: (2026)
Effects of auditory distance cues and reverberation on spatial perception and listening strategies
by: Missoni, Fulvio, et al.
Published: (2025)
by: Missoni, Fulvio, et al.
Published: (2025)
LLM supervised Pre-training for Multimodal Emotion Recognition in Conversations
by: Dutta, Soumya, et al.
Published: (2025)
by: Dutta, Soumya, et al.
Published: (2025)
ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals
by: E, Ameenudeen P, et al.
Published: (2026)
by: E, Ameenudeen P, et al.
Published: (2026)
Towards predicting binaural audio quality in listeners with normal and impaired hearing
by: Biberger, Thomas, et al.
Published: (2025)
by: Biberger, Thomas, et al.
Published: (2025)
SuperM2M: Supervised and Mixture-to-Mixture Co-Learning for Speech Enhancement and Noise-Robust ASR
by: Wang, Zhong-Qiu
Published: (2024)
by: Wang, Zhong-Qiu
Published: (2024)
Unsupervised Face-Masked Speech Enhancement Using Generative Adversarial Networks With Human-in-the-Loop Assessment Metrics
by: Wang, Syu-Siang, et al.
Published: (2024)
by: Wang, Syu-Siang, et al.
Published: (2024)
Efficient Personalization of Amplification in Hearing Aids via Multi-band Bayesian Machine Learning
by: Ni, Aoxin, et al.
Published: (2024)
by: Ni, Aoxin, et al.
Published: (2024)
ERes2NetV2: Boosting Short-Duration Speaker Verification Performance with Computational Efficiency
by: Chen, Yafeng, et al.
Published: (2024)
by: Chen, Yafeng, et al.
Published: (2024)
Classification of Adventitious Sounds Combining Cochleogram and Vision Transformers
by: Mang, Loredana Daria, et al.
Published: (2024)
by: Mang, Loredana Daria, et al.
Published: (2024)
Advanced Signal Analysis in Detecting Replay Attacks for Automatic Speaker Verification Systems
by: Kuang, Lee Shih
Published: (2024)
by: Kuang, Lee Shih
Published: (2024)
Cough-E: A multimodal, privacy-preserving cough detection algorithm for the edge
by: Albini, Stefano, et al.
Published: (2024)
by: Albini, Stefano, et al.
Published: (2024)
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
by: Damiano, Stefano, et al.
Published: (2024)
by: Damiano, Stefano, et al.
Published: (2024)
A Neural Denoising Vocoder for Clean Waveform Generation from Noisy Mel-Spectrogram based on Amplitude and Phase Predictions
by: Du, Hui-Peng, et al.
Published: (2024)
by: Du, Hui-Peng, et al.
Published: (2024)
BanglaNum -- A Public Dataset for Bengali Digit Recognition from Speech
by: Mohammad, Mir Sayeed, et al.
Published: (2024)
by: Mohammad, Mir Sayeed, et al.
Published: (2024)
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
by: Fan, Shitong, et al.
Published: (2024)
by: Fan, Shitong, et al.
Published: (2024)
Unsupervised detection and classification of heartbeats using the dissimilarity matrix in PCG signals
by: Torre-Cruz, J., et al.
Published: (2024)
by: Torre-Cruz, J., et al.
Published: (2024)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
by: Miotello, Federico, et al.
Published: (2024)
by: Miotello, Federico, et al.
Published: (2024)
Speaker and Style Disentanglement of Speech Based on Contrastive Predictive Coding Supported Factorized Variational Autoencoder
by: Xie, Yuying, et al.
Published: (2024)
by: Xie, Yuying, et al.
Published: (2024)
DJ Mix Transcription with Multi-Pass Non-Negative Matrix Factorization
by: André, Étienne Paul, et al.
Published: (2024)
by: André, Étienne Paul, et al.
Published: (2024)
Robust Fixed-Filter Sound Zone Control with Audio-Based Position Tracking
by: Bhattacharjee, Sankha Subhra, et al.
Published: (2024)
by: Bhattacharjee, Sankha Subhra, et al.
Published: (2024)
3D-Speaker-Toolkit: An Open-Source Toolkit for Multimodal Speaker Verification and Diarization
by: Chen, Yafeng, et al.
Published: (2024)
by: Chen, Yafeng, et al.
Published: (2024)
PLDNet: PLD-Guided Lightweight Deep Network Boosted by Efficient Attention for Handheld Dual-Microphone Speech Enhancement
by: Zhou, Nan, et al.
Published: (2024)
by: Zhou, Nan, et al.
Published: (2024)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
by: MacWilliam, Kathleen, et al.
Published: (2024)
by: MacWilliam, Kathleen, et al.
Published: (2024)
Detection of manatee vocalisations using the Audio Spectrogram Transformer
by: Schiappacasse, Stefano, et al.
Published: (2024)
by: Schiappacasse, Stefano, et al.
Published: (2024)
A Novel Numerical Method for Relaxing the Minimal Configurations of TOA-Based Joint Sensors and Sources Localization
by: Cao, Faxian, et al.
Published: (2024)
by: Cao, Faxian, et al.
Published: (2024)
Zero-Bit Transmission of Adaptive Pre- and De-emphasis Filters for Speech and Audio Coding
by: Piralideh, Niloofar Omidi, et al.
Published: (2024)
by: Piralideh, Niloofar Omidi, et al.
Published: (2024)
Mixture to Mixture: Leveraging Close-talk Mixtures as Weak-supervision for Speech Separation
by: Wang, Zhong-Qiu
Published: (2024)
by: Wang, Zhong-Qiu
Published: (2024)
USDnet: Unsupervised Speech Dereverberation via Neural Forward Filtering
by: Wang, Zhong-Qiu
Published: (2024)
by: Wang, Zhong-Qiu
Published: (2024)
Revisiting proximity effect using broadband signals
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation
by: Kim, Miseul, et al.
Published: (2024)
by: Kim, Miseul, et al.
Published: (2024)
Similar Items
-
Benchmarking Humans and Machines on Complex Multilingual Speech Understanding Tasks
by: Kankanala, Sai Samrat, et al.
Published: (2025) -
Summary of the DISPLACE Challenge 2023 - DIarization of SPeaker and LAnguage in Conversational Environments
by: Baghel, Shikha, et al.
Published: (2023) -
On fusing active and passive acoustic sensing for simultaneous localization and mapping
by: Bradley, Aidan J., et al.
Published: (2024) -
improving partition-block-based acoustic echo canceler in under-modeling scenarios
by: Fan, Wenzhi, et al.
Published: (2020) -
RIR-Mega: a large-scale simulated room impulse response dataset for machine learning and room acoustics modeling
by: Goswami, Mandip
Published: (2025)