Gespeichert in:
| Hauptverfasser: | Sonck, Rien, Accou, Bernd, Francart, Tom, Vanthornhout, Jonas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.03624 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
von: Geirnaert, Simon, et al.
Veröffentlicht: (2025)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2025)
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
von: De Clercq, Pieter, et al.
Veröffentlicht: (2024)
von: De Clercq, Pieter, et al.
Veröffentlicht: (2024)
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis
von: Heintz, Nicolas, et al.
Veröffentlicht: (2025)
von: Heintz, Nicolas, et al.
Veröffentlicht: (2025)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
CLASH: Contrastive learning through alignment shifting to extract stimulus information from EEG
von: Accou, Bernd, et al.
Veröffentlicht: (2023)
von: Accou, Bernd, et al.
Veröffentlicht: (2023)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2026)
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2026)
On Improving Error Resilience of Neural End-to-End Speech Coders
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
Frequency-Specific Neural Response and Cross-Correlation Analysis of Envelope Following Responses to Native Speech and Music Using Multichannel EEG Signals: A Case Study
von: Hasan, Md. Mahbub, et al.
Veröffentlicht: (2025)
von: Hasan, Md. Mahbub, et al.
Veröffentlicht: (2025)
A DNN Based Post-Filter to Enhance the Quality of Coded Speech in MDCT Domain
von: Gupta, Kishan, et al.
Veröffentlicht: (2022)
von: Gupta, Kishan, et al.
Veröffentlicht: (2022)
A Data-Centric Approach to Generalizable Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
A Robust Method for Pitch Tracking in the Frequency Following Response using Harmonic Amplitude Summation Filterbank
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners
von: Yuan, Ze, et al.
Veröffentlicht: (2024)
von: Yuan, Ze, et al.
Veröffentlicht: (2024)
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
Audio-Visual Speech Enhancement: Architectural Design and Deployment Strategies
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
Latent Granular Resynthesis using Neural Audio Codecs
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
von: Wilroth, Johanna, et al.
Veröffentlicht: (2026)
von: Wilroth, Johanna, et al.
Veröffentlicht: (2026)
QINCODEC: Neural Audio Compression with Implicit Neural Codebooks
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2025)
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2025)
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
ASVspoof 5: Evaluation of Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
von: Wang, Xin, et al.
Veröffentlicht: (2026)
von: Wang, Xin, et al.
Veröffentlicht: (2026)
A Study on Speech Assessment with Visual Cues
von: Ahmed, Shafique, et al.
Veröffentlicht: (2025)
von: Ahmed, Shafique, et al.
Veröffentlicht: (2025)
Semantic Communications for Speech Recognition
von: Weng, Zhenzi, et al.
Veröffentlicht: (2021)
von: Weng, Zhenzi, et al.
Veröffentlicht: (2021)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
von: Abu, Avi, et al.
Veröffentlicht: (2024)
von: Abu, Avi, et al.
Veröffentlicht: (2024)
Stimulus-Informed Generalized Canonical Correlation Analysis for Group Analysis of Neural Responses to Natural Stimuli
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Joint Semantic Knowledge Distillation and Masked Acoustic Modeling for Full-band Speech Restoration with Improved Intelligibility
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
von: Kim, Minje, et al.
Veröffentlicht: (2024)
von: Kim, Minje, et al.
Veröffentlicht: (2024)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
von: Iatariene, Taous, et al.
Veröffentlicht: (2025)
von: Iatariene, Taous, et al.
Veröffentlicht: (2025)
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Investigation of Feature Selection and Pooling Methods for Environmental Sound Classification
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
Fully Reversing the Shoebox Image Source Method: From Impulse Responses to Room Parameters
von: Sprunck, Tom, et al.
Veröffentlicht: (2024)
von: Sprunck, Tom, et al.
Veröffentlicht: (2024)
Toward Universal Speech Enhancement for Diverse Input Conditions
von: Zhang, Wangyou, et al.
Veröffentlicht: (2023)
von: Zhang, Wangyou, et al.
Veröffentlicht: (2023)
Speech dereverberation constrained on room impulse response characteristics
von: Bahrman, Louis, et al.
Veröffentlicht: (2024)
von: Bahrman, Louis, et al.
Veröffentlicht: (2024)
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
Predicting Heart Activity from Speech using Data-driven and Knowledge-based features
von: Elbanna, Gasser, et al.
Veröffentlicht: (2024)
von: Elbanna, Gasser, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
von: Geirnaert, Simon, et al.
Veröffentlicht: (2025) -
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
von: De Clercq, Pieter, et al.
Veröffentlicht: (2024) -
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis
von: Heintz, Nicolas, et al.
Veröffentlicht: (2025) -
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024) -
CLASH: Contrastive learning through alignment shifting to extract stimulus information from EEG
von: Accou, Bernd, et al.
Veröffentlicht: (2023)