The Unreliability of Acoustic Systems in Alzheimer's Speech Datasets with Heterogeneous Recording Conditions
Fuente:
arXiv
Salvato in:
| Autori principali: | Gauder, Lara, Riera, Pablo, Slachevsky, Andrea, Forno, Gonzalo, Garcia, Adolfo M., Ferrer, Luciana |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EnCodecMAE: Leveraging neural codecs for universal audio representation learning
di: Pepino, Leonardo, et al.
Pubblicazione: (2023)
di: Pepino, Leonardo, et al.
Pubblicazione: (2023)
A Toolkit for Detecting Spurious Correlations in Speech Datasets
di: Gauder, Lara, et al.
Pubblicazione: (2026)
di: Gauder, Lara, et al.
Pubblicazione: (2026)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
di: Pepino, Leonardo, et al.
Pubblicazione: (2024)
di: Pepino, Leonardo, et al.
Pubblicazione: (2024)
Study on the Fairness of Speaker Verification Systems on Underrepresented Accents in English
di: Estevez, Mariel, et al.
Pubblicazione: (2022)
di: Estevez, Mariel, et al.
Pubblicazione: (2022)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
di: Yang, Bing, et al.
Pubblicazione: (2024)
di: Yang, Bing, et al.
Pubblicazione: (2024)
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
di: Gao, Yuan, et al.
Pubblicazione: (2025)
di: Gao, Yuan, et al.
Pubblicazione: (2025)
Acoustic BPE for Speech Generation with Discrete Tokens
di: Shen, Feiyu, et al.
Pubblicazione: (2023)
di: Shen, Feiyu, et al.
Pubblicazione: (2023)
Robust Audio Tagging under Class-wise Supervision Unreliability
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
di: Hou, Yuanbo, et al.
Pubblicazione: (2026)
Abusive Speech Detection in Indic Languages Using Acoustic Features
di: Spiesberger, Anika A., et al.
Pubblicazione: (2024)
di: Spiesberger, Anika A., et al.
Pubblicazione: (2024)
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
di: Garg, Ashi, et al.
Pubblicazione: (2025)
di: Garg, Ashi, et al.
Pubblicazione: (2025)
Improving Acoustic Scene Classification in Low-Resource Conditions
di: Chen, Zhi, et al.
Pubblicazione: (2024)
di: Chen, Zhi, et al.
Pubblicazione: (2024)
Room Impulse Response Generation Conditioned on Acoustic Parameters
di: Arellano, Silvia, et al.
Pubblicazione: (2025)
di: Arellano, Silvia, et al.
Pubblicazione: (2025)
DIFFRENT: A Diffusion Model for Recording Environment Transfer of Speech
di: Im, Jaekwon, et al.
Pubblicazione: (2024)
di: Im, Jaekwon, et al.
Pubblicazione: (2024)
Neural Speech Tracking in a Virtual Acoustic Environment: Audio-Visual Benefit for Unscripted Continuous Speech
di: Daeglau, Mareike, et al.
Pubblicazione: (2025)
di: Daeglau, Mareike, et al.
Pubblicazione: (2025)
From Human Speech to Ocean Signals: Transferring Speech Large Models for Underwater Acoustic Target Recognition
di: Huang, Mengcheng, et al.
Pubblicazione: (2026)
di: Huang, Mengcheng, et al.
Pubblicazione: (2026)
Temporally Heterogeneous Graph Contrastive Learning for Multimodal Acoustic event Classification
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis
di: Yang, Qian, et al.
Pubblicazione: (2024)
di: Yang, Qian, et al.
Pubblicazione: (2024)
SAC: Neural Speech Codec with Semantic-Acoustic Dual-Stream Quantization
di: Chen, Wenxi, et al.
Pubblicazione: (2025)
di: Chen, Wenxi, et al.
Pubblicazione: (2025)
Comparative Evaluation of Acoustic Feature Extraction Tools for Clinical Speech Analysis
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2025)
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2025)
UrBAN: Urban Beehive Acoustics and PheNotyping Dataset
di: Abdollahi, Mahsa, et al.
Pubblicazione: (2024)
di: Abdollahi, Mahsa, et al.
Pubblicazione: (2024)
VoiceRestore: Flow-Matching Transformers for Speech Recording Quality Restoration
di: Kirdey, Stanislav
Pubblicazione: (2025)
di: Kirdey, Stanislav
Pubblicazione: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
di: Chung, Soo-Whan, et al.
Pubblicazione: (2025)
di: Chung, Soo-Whan, et al.
Pubblicazione: (2025)
Evaluating Parkinson's Disease Detection in Anonymized Speech: A Performance and Acoustic Analysis
di: Franzreb, Carlos, et al.
Pubblicazione: (2026)
di: Franzreb, Carlos, et al.
Pubblicazione: (2026)
Continuous Target Speech Extraction: Enhancing Personalized Diarization and Extraction on Complex Recordings
di: Zhao, He, et al.
Pubblicazione: (2024)
di: Zhao, He, et al.
Pubblicazione: (2024)
DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions
di: Chen, Weidong, et al.
Pubblicazione: (2025)
di: Chen, Weidong, et al.
Pubblicazione: (2025)
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature
di: Du, Chenpeng, et al.
Pubblicazione: (2022)
di: Du, Chenpeng, et al.
Pubblicazione: (2022)
Benchmarking Time-localized Explanations for Audio Classification Models
di: Bolaños, Cecilia, et al.
Pubblicazione: (2025)
di: Bolaños, Cecilia, et al.
Pubblicazione: (2025)
SEABAD: A Tropical Bird Activity Detection Dataset for Passive Acoustic Monitoring
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
Leveraging Multimodal Methods and Spontaneous Speech for Alzheimer's Disease Identification
di: Gao, Yifan, et al.
Pubblicazione: (2024)
di: Gao, Yifan, et al.
Pubblicazione: (2024)
Parallel GPT: Harmonizing the Independence and Interdependence of Acoustic and Semantic Information for Zero-Shot Text-to-Speech
di: Xing, Jingyuan, et al.
Pubblicazione: (2025)
di: Xing, Jingyuan, et al.
Pubblicazione: (2025)
Improving Design of Input Condition Invariant Speech Enhancement
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
CUEMPATHY: A Counseling Speech Dataset for Psychotherapy Research
di: Tao, Dehua, et al.
Pubblicazione: (2024)
di: Tao, Dehua, et al.
Pubblicazione: (2024)
Dataset-Distillation Generative Model for Speech Emotion Recognition
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2024)
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2024)
Confidence-based Filtering for Speech Dataset Curation with Generative Speech Enhancement Using Discrete Tokens
di: Yamauchi, Kazuki, et al.
Pubblicazione: (2026)
di: Yamauchi, Kazuki, et al.
Pubblicazione: (2026)
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
di: Tailleur, Modan, et al.
Pubblicazione: (2025)
di: Tailleur, Modan, et al.
Pubblicazione: (2025)
Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition
di: Xie, Xurong, et al.
Pubblicazione: (2022)
di: Xie, Xurong, et al.
Pubblicazione: (2022)
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards
di: Ren, Yong, et al.
Pubblicazione: (2026)
di: Ren, Yong, et al.
Pubblicazione: (2026)
Improving Language Model-Based Zero-Shot Text-to-Speech Synthesis with Multi-Scale Acoustic Prompts
di: Lei, Shun, et al.
Pubblicazione: (2023)
di: Lei, Shun, et al.
Pubblicazione: (2023)
Effects of Recording Condition and Number of Monitored Days on Discriminative Power of the Daily Phonotrauma Index
di: Ghasemzadeh, Hamzeh, et al.
Pubblicazione: (2024)
di: Ghasemzadeh, Hamzeh, et al.
Pubblicazione: (2024)
LipDiffuser: Lip-to-Speech Generation with Conditional Diffusion Models
di: Richter, Julius, et al.
Pubblicazione: (2025)
di: Richter, Julius, et al.
Pubblicazione: (2025)
Documenti analoghi
-
EnCodecMAE: Leveraging neural codecs for universal audio representation learning
di: Pepino, Leonardo, et al.
Pubblicazione: (2023) -
A Toolkit for Detecting Spurious Correlations in Speech Datasets
di: Gauder, Lara, et al.
Pubblicazione: (2026) -
Fusion approaches for emotion recognition from speech using acoustic and text-based features
di: Pepino, Leonardo, et al.
Pubblicazione: (2024) -
Study on the Fairness of Speaker Verification Systems on Underrepresented Accents in English
di: Estevez, Mariel, et al.
Pubblicazione: (2022) -
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
di: Yang, Bing, et al.
Pubblicazione: (2024)