A Multi-decoder Neural Tracking Method for Accurately Predicting Speech Intelligibility
Fuente:
arXiv
Salvato in:
| Autori principali: | Sonck, Rien, Accou, Bernd, Francart, Tom, Vanthornhout, Jonas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
di: Geirnaert, Simon, et al.
Pubblicazione: (2025)
di: Geirnaert, Simon, et al.
Pubblicazione: (2025)
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis
di: Heintz, Nicolas, et al.
Pubblicazione: (2025)
di: Heintz, Nicolas, et al.
Pubblicazione: (2025)
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
di: De Clercq, Pieter, et al.
Pubblicazione: (2024)
di: De Clercq, Pieter, et al.
Pubblicazione: (2024)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
di: Geirnaert, Simon, et al.
Pubblicazione: (2024)
di: Geirnaert, Simon, et al.
Pubblicazione: (2024)
CLASH: Contrastive learning through alignment shifting to extract stimulus information from EEG
di: Accou, Bernd, et al.
Pubblicazione: (2023)
di: Accou, Bernd, et al.
Pubblicazione: (2023)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
di: Maghsoudi, Maryam, et al.
Pubblicazione: (2026)
di: Maghsoudi, Maryam, et al.
Pubblicazione: (2026)
A Data-Centric Approach to Generalizable Speech Deepfake Detection
di: Huang, Wen, et al.
Pubblicazione: (2025)
di: Huang, Wen, et al.
Pubblicazione: (2025)
On Improving Error Resilience of Neural End-to-End Speech Coders
di: Gupta, Kishan, et al.
Pubblicazione: (2024)
di: Gupta, Kishan, et al.
Pubblicazione: (2024)
A Robust Method for Pitch Tracking in the Frequency Following Response using Harmonic Amplitude Summation Filterbank
di: Sadeghkhani, Sajad, et al.
Pubblicazione: (2025)
di: Sadeghkhani, Sajad, et al.
Pubblicazione: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
di: Rahimi, Akam, et al.
Pubblicazione: (2025)
Audio-Visual Speech Enhancement: Architectural Design and Deployment Strategies
di: Hamadouche, Anis, et al.
Pubblicazione: (2025)
di: Hamadouche, Anis, et al.
Pubblicazione: (2025)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners
di: Yuan, Ze, et al.
Pubblicazione: (2024)
di: Yuan, Ze, et al.
Pubblicazione: (2024)
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
di: Lu, Ye-Xin, et al.
Pubblicazione: (2024)
di: Lu, Ye-Xin, et al.
Pubblicazione: (2024)
QINCODEC: Neural Audio Compression with Implicit Neural Codebooks
di: Lahrichi, Zineb, et al.
Pubblicazione: (2025)
di: Lahrichi, Zineb, et al.
Pubblicazione: (2025)
A DNN Based Post-Filter to Enhance the Quality of Coded Speech in MDCT Domain
di: Gupta, Kishan, et al.
Pubblicazione: (2022)
di: Gupta, Kishan, et al.
Pubblicazione: (2022)
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
di: Thakkar, Karan, et al.
Pubblicazione: (2026)
di: Thakkar, Karan, et al.
Pubblicazione: (2026)
ASVspoof 5: Evaluation of Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
di: Wang, Xin, et al.
Pubblicazione: (2026)
di: Wang, Xin, et al.
Pubblicazione: (2026)
Frequency-Specific Neural Response and Cross-Correlation Analysis of Envelope Following Responses to Native Speech and Music Using Multichannel EEG Signals: A Case Study
di: Hasan, Md. Mahbub, et al.
Pubblicazione: (2025)
di: Hasan, Md. Mahbub, et al.
Pubblicazione: (2025)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
di: Wilroth, Johanna, et al.
Pubblicazione: (2026)
di: Wilroth, Johanna, et al.
Pubblicazione: (2026)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
di: Qi, Tianhua, et al.
Pubblicazione: (2026)
di: Qi, Tianhua, et al.
Pubblicazione: (2026)
A Study on Speech Assessment with Visual Cues
di: Ahmed, Shafique, et al.
Pubblicazione: (2025)
di: Ahmed, Shafique, et al.
Pubblicazione: (2025)
Latent Granular Resynthesis using Neural Audio Codecs
di: Tokui, Nao, et al.
Pubblicazione: (2025)
di: Tokui, Nao, et al.
Pubblicazione: (2025)
Investigation of Feature Selection and Pooling Methods for Environmental Sound Classification
di: Dehaghani, Parinaz Binandeh, et al.
Pubblicazione: (2025)
di: Dehaghani, Parinaz Binandeh, et al.
Pubblicazione: (2025)
Semantic Communications for Speech Recognition
di: Weng, Zhenzi, et al.
Pubblicazione: (2021)
di: Weng, Zhenzi, et al.
Pubblicazione: (2021)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
di: Abu, Avi, et al.
Pubblicazione: (2024)
di: Abu, Avi, et al.
Pubblicazione: (2024)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
di: Iatariene, Taous, et al.
Pubblicazione: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
di: Lu, Jiajun, et al.
Pubblicazione: (2023)
di: Lu, Jiajun, et al.
Pubblicazione: (2023)
Toward Universal Speech Enhancement for Diverse Input Conditions
di: Zhang, Wangyou, et al.
Pubblicazione: (2023)
di: Zhang, Wangyou, et al.
Pubblicazione: (2023)
Speech dereverberation constrained on room impulse response characteristics
di: Bahrman, Louis, et al.
Pubblicazione: (2024)
di: Bahrman, Louis, et al.
Pubblicazione: (2024)
Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
di: Yang, Yujie, et al.
Pubblicazione: (2025)
di: Yang, Yujie, et al.
Pubblicazione: (2025)
Joint Semantic Knowledge Distillation and Masked Acoustic Modeling for Full-band Speech Restoration with Improved Intelligibility
di: Liu, Xiaoyu, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyu, et al.
Pubblicazione: (2024)
Stimulus-Informed Generalized Canonical Correlation Analysis for Group Analysis of Neural Responses to Natural Stimuli
di: Geirnaert, Simon, et al.
Pubblicazione: (2024)
di: Geirnaert, Simon, et al.
Pubblicazione: (2024)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
di: Kim, Minje, et al.
Pubblicazione: (2024)
di: Kim, Minje, et al.
Pubblicazione: (2024)
Modulation Feature Enhancement with a Multi-Stage Attention Network for Underwater Acoustic Target Recognition
di: Yu, Jiaping, et al.
Pubblicazione: (2026)
di: Yu, Jiaping, et al.
Pubblicazione: (2026)
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
di: Neri, Michael, et al.
Pubblicazione: (2025)
di: Neri, Michael, et al.
Pubblicazione: (2025)
Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement
di: Serre, Thomas, et al.
Pubblicazione: (2026)
di: Serre, Thomas, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
di: Geirnaert, Simon, et al.
Pubblicazione: (2025) -
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis
di: Heintz, Nicolas, et al.
Pubblicazione: (2025) -
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
di: De Clercq, Pieter, et al.
Pubblicazione: (2024) -
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
di: Geirnaert, Simon, et al.
Pubblicazione: (2024) -
CLASH: Contrastive learning through alignment shifting to extract stimulus information from EEG
di: Accou, Bernd, et al.
Pubblicazione: (2023)