XANE Background Acoustic Embeddings: Ablation and Clustering Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Dushyant, Fosburgh, James, Dumpala, Sri Harsha, Sastri, Chandramouli Shama, Kruchinin, Stanislav Yu., Naylor, Patrick A. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
XANE: eXplainable Acoustic Neural Embeddings
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Test-Time Training for Depression Detection
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Self-Supervised Embeddings for Detecting Individual Symptoms of Depression
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Predicting Individual Depression Symptoms from Acoustic Features During Speech
by: Rodriguez, Sebastian, et al.
Published: (2024)
by: Rodriguez, Sebastian, et al.
Published: (2024)
Categorical Unsupervised Variational Acoustic Clustering
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
Unsupervised Variational Acoustic Clustering
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
Clustering of Acoustic Environments with Variational Autoencoders for Hearing Devices
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
by: Fiorio, Luan Vinícius, et al.
Published: (2025)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
by: Nespoli, Francesco, et al.
Published: (2024)
by: Nespoli, Francesco, et al.
Published: (2024)
Layer-Aware Early Fusion of Acoustic and Linguistic Embeddings for Cognitive Status Classification
by: Novotny, Krystof, et al.
Published: (2026)
by: Novotny, Krystof, et al.
Published: (2026)
Matching Reverberant Speech Through Learned Acoustic Embeddings and Feedback Delay Networks
by: Götz, Philipp, et al.
Published: (2025)
by: Götz, Philipp, et al.
Published: (2025)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
by: Götz, Philipp, et al.
Published: (2024)
by: Götz, Philipp, et al.
Published: (2024)
Binaural Speech Enhancement Using Deep Complex Convolutional Transformer Networks
by: Tokala, Vikas, et al.
Published: (2024)
by: Tokala, Vikas, et al.
Published: (2024)
Head-steered channel selection method for hearing aid applications using remote microphones
by: Sathyapriyan, Vasudha, et al.
Published: (2025)
by: Sathyapriyan, Vasudha, et al.
Published: (2025)
Sub-band Domain Multi-Hypothesis Acoustic Echo Canceler Based Acoustic Scene Analysis
by: Southwell, Benjamin J, et al.
Published: (2025)
by: Southwell, Benjamin J, et al.
Published: (2025)
The Neural-SRP method for positional sound source localization
by: Grinstein, Eric, et al.
Published: (2024)
by: Grinstein, Eric, et al.
Published: (2024)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
by: Jiang, Hao, et al.
Published: (2026)
by: Jiang, Hao, et al.
Published: (2026)
Embedded Acoustic Intelligence for Automotive Systems
by: Rajagopal, Renjith, et al.
Published: (2025)
by: Rajagopal, Renjith, et al.
Published: (2025)
Acoustic-to-articulatory Inversion of the Complete Vocal Tract from RT-MRI with Various Audio Embeddings and Dataset Sizes
by: Azzouz, Sofiane, et al.
Published: (2026)
by: Azzouz, Sofiane, et al.
Published: (2026)
Spatio-temporal Latent Representations for the Analysis of Acoustic Scenes in-the-wild
by: Montero-Ramírez, Claudia, et al.
Published: (2024)
by: Montero-Ramírez, Claudia, et al.
Published: (2024)
Comparison of Tiny Machine Learning Techniques for Embedded Acoustic Emission Analysis
by: Muthumala, Uditha, et al.
Published: (2024)
by: Muthumala, Uditha, et al.
Published: (2024)
Perceptual evaluation of Acoustic Level of Detail in Virtual Acoustic Environments
by: Fichna, Stefan, et al.
Published: (2025)
by: Fichna, Stefan, et al.
Published: (2025)
Binaural Speech Enhancement Using Complex Convolutional Recurrent Networks
by: Tokala, Vikas, et al.
Published: (2025)
by: Tokala, Vikas, et al.
Published: (2025)
Uncertainty Quantification in Machine Learning for Joint Speaker Diarization and Identification
by: McKnight, Simon W., et al.
Published: (2023)
by: McKnight, Simon W., et al.
Published: (2023)
Emotional Styles Hide in Deep Speaker Embeddings: Disentangle Deep Speaker Embeddings for Speaker Clustering
by: Lin, Chaohao, et al.
Published: (2025)
by: Lin, Chaohao, et al.
Published: (2025)
Binaural Localization Model for Speech in Noise
by: Tokala, Vikas, et al.
Published: (2025)
by: Tokala, Vikas, et al.
Published: (2025)
What Does the Speaker Embedding Encode?
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
On the Use of Dereverberation for Acoustic Feedback Cancellation
by: Liekens, Basil, et al.
Published: (2026)
by: Liekens, Basil, et al.
Published: (2026)
Scene-wide Acoustic Parameter Estimation
by: Falcon-Perez, Ricardo, et al.
Published: (2024)
by: Falcon-Perez, Ricardo, et al.
Published: (2024)
VoiceRestore: Flow-Matching Transformers for Speech Recording Quality Restoration
by: Kirdey, Stanislav
Published: (2025)
by: Kirdey, Stanislav
Published: (2025)
Acoustic and Semantic Modeling of Emotion in Spoken Language
by: Dutta, Soumya
Published: (2026)
by: Dutta, Soumya
Published: (2026)
A Theoretical Framework for Acoustic Neighbor Embeddings
by: Jeon, Woojay
Published: (2024)
by: Jeon, Woojay
Published: (2024)
The Reasonable Effectiveness of Speaker Embeddings for Violence Detection
by: Jain, Sarthak, et al.
Published: (2024)
by: Jain, Sarthak, et al.
Published: (2024)
Decoding Vocal Articulations from Acoustic Latent Representations
by: Cámara, Mateo, et al.
Published: (2024)
by: Cámara, Mateo, et al.
Published: (2024)
Acoustic to Articulatory Speech Inversion for Children with Velopharyngeal Insufficiency
by: Tabatabaee, Saba, et al.
Published: (2025)
by: Tabatabaee, Saba, et al.
Published: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
by: Dutta, Soumya, et al.
Published: (2024)
by: Dutta, Soumya, et al.
Published: (2024)
Enhancing Acoustic-to-Articulatory Speech Inversion by Incorporating Nasality
by: Tabatabaee, Saba, et al.
Published: (2025)
by: Tabatabaee, Saba, et al.
Published: (2025)
Comparative Evaluation of Acoustic Feature Extraction Tools for Clinical Speech Analysis
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
by: Choi, Anna Seo Gyeong, et al.
Published: (2025)
Acoustic BPE for Speech Generation with Discrete Tokens
by: Shen, Feiyu, et al.
Published: (2023)
by: Shen, Feiyu, et al.
Published: (2023)
Evaluation of Virtual Acoustic Environments with Different Acoustic Level of Detail
by: Fichna, Stefan, et al.
Published: (2023)
by: Fichna, Stefan, et al.
Published: (2023)
SAGA-SR: Semantically and Acoustically Guided Audio Super-Resolution
by: Im, Jaekwon, et al.
Published: (2025)
by: Im, Jaekwon, et al.
Published: (2025)
Similar Items
-
XANE: eXplainable Acoustic Neural Embeddings
by: Dumpala, Sri Harsha, et al.
Published: (2024) -
Test-Time Training for Depression Detection
by: Dumpala, Sri Harsha, et al.
Published: (2024) -
Self-Supervised Embeddings for Detecting Individual Symptoms of Depression
by: Dumpala, Sri Harsha, et al.
Published: (2024) -
Predicting Individual Depression Symptoms from Acoustic Features During Speech
by: Rodriguez, Sebastian, et al.
Published: (2024) -
Categorical Unsupervised Variational Acoustic Clustering
by: Fiorio, Luan Vinícius, et al.
Published: (2025)