Comparison of linear and nonlinear methods for decoding selective attention to speech from ear-EEG recordings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thornton, Mike, Mandic, Danilo, Reichenbach, Tobias |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
von: Thornton, Mike, et al.
Veröffentlicht: (2024)
Joint decoding method for controllable contextual speech recognition based on Speech LLM
von: Fang, Yangui, et al.
Veröffentlicht: (2025)
von: Fang, Yangui, et al.
Veröffentlicht: (2025)
Explainable speech emotion recognition through attentive pooling: insights from attention-based temporal localization
von: Leygue, Tahitoa, et al.
Veröffentlicht: (2025)
von: Leygue, Tahitoa, et al.
Veröffentlicht: (2025)
Omni-directional attention mechanism based on Mamba for speech separation
von: Xue, Ke, et al.
Veröffentlicht: (2026)
von: Xue, Ke, et al.
Veröffentlicht: (2026)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Head-steered channel selection method for hearing aid applications using remote microphones
von: Sathyapriyan, Vasudha, et al.
Veröffentlicht: (2025)
von: Sathyapriyan, Vasudha, et al.
Veröffentlicht: (2025)
Synthesizing speech with selected perceptual voice qualities - A case study with creaky voice
von: Rautenberg, Frederik, et al.
Veröffentlicht: (2025)
von: Rautenberg, Frederik, et al.
Veröffentlicht: (2025)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Heterogeneous bimodal attention fusion for speech emotion recognition
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
Exploring the limits of decoder-only models trained on public speech recognition corpora
von: Gupta, Ankit, et al.
Veröffentlicht: (2024)
von: Gupta, Ankit, et al.
Veröffentlicht: (2024)
Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
von: Shirahata, Yuma, et al.
Veröffentlicht: (2024)
von: Shirahata, Yuma, et al.
Veröffentlicht: (2024)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
von: Saon, George, et al.
Veröffentlicht: (2025)
von: Saon, George, et al.
Veröffentlicht: (2025)
SLM-S2ST: A multimodal language model for direct speech-to-speech translation
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
On combining acoustic and modulation spectrograms in an attention LSTM-based system for speech intelligibility level classification
von: Gallardo-Antolín, Ascensión, et al.
Veröffentlicht: (2024)
von: Gallardo-Antolín, Ascensión, et al.
Veröffentlicht: (2024)
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions
von: Mack, Wolfgang, et al.
Veröffentlicht: (2025)
von: Mack, Wolfgang, et al.
Veröffentlicht: (2025)
Graph-based multi-Feature fusion method for speech emotion recognition
von: Liu, Xueyu, et al.
Veröffentlicht: (2024)
von: Liu, Xueyu, et al.
Veröffentlicht: (2024)
A lightweight and robust method for blind wideband-to-fullband extension of speech
von: Büthe, Jan, et al.
Veröffentlicht: (2024)
von: Büthe, Jan, et al.
Veröffentlicht: (2024)
Predicting speech intelligibility in older adults for speech enhancement using the Gammachirp Envelope Similarity Index, GESI
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2025)
Good practices for evaluation of synthesized speech
von: Cooper, Erica, et al.
Veröffentlicht: (2025)
von: Cooper, Erica, et al.
Veröffentlicht: (2025)
On the reliability of feature attribution methods for speech classification
von: Shen, Gaofei, et al.
Veröffentlicht: (2025)
von: Shen, Gaofei, et al.
Veröffentlicht: (2025)
Eardrum sound pressure prediction from ear canal reflectance based on the inverse solution of Webster's horn equation
von: Roden, Reinhild, et al.
Veröffentlicht: (2025)
von: Roden, Reinhild, et al.
Veröffentlicht: (2025)
RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
LDCodec: A high quality neural audio codec with low-complexity decoder
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
How to train your ears: Auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing losses
von: Leer, Peter, et al.
Veröffentlicht: (2024)
von: Leer, Peter, et al.
Veröffentlicht: (2024)
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2026)
von: Lee, Dongheon, et al.
Veröffentlicht: (2026)
Short-term cognitive fatigue of spatial selective attention after face-to-face conversations in virtual noisy environments
von: Hládek, Ľuboš, et al.
Veröffentlicht: (2025)
von: Hládek, Ľuboš, et al.
Veröffentlicht: (2025)
Non-invasive electromyographic speech neuroprosthesis: a geometric perspective
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
von: Gowda, Harshavardhana T., et al.
Veröffentlicht: (2025)
Modeling of Speech-dependent Own Voice Transfer Characteristics for Hearables with In-ear Microphones
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
Speech-dependent Modeling of Own Voice Transfer Characteristics for In-ear Microphones in Hearables
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
von: Ohlenbusch, Mattes, et al.
Veröffentlicht: (2023)
BabAR: from phoneme recognition to developmental measures of young children's speech production
von: Lavechin, Marvin, et al.
Veröffentlicht: (2026)
von: Lavechin, Marvin, et al.
Veröffentlicht: (2026)
Token-based Attractors and Cross-attention in Spoof Diarization
von: Koo, Kyo-Won, et al.
Veröffentlicht: (2025)
von: Koo, Kyo-Won, et al.
Veröffentlicht: (2025)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
von: Deng, Qingkun, et al.
Veröffentlicht: (2024)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
von: Tabatabaee, Saba, et al.
Veröffentlicht: (2026)
Transcribe, Align and Segment: Creating speech datasets for low-resource languages
von: Sereda, Taras
Veröffentlicht: (2024)
von: Sereda, Taras
Veröffentlicht: (2024)
Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation
von: Ma, Jianbo, et al.
Veröffentlicht: (2026)
von: Ma, Jianbo, et al.
Veröffentlicht: (2026)
On the relationship between speech and hearing
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
von: Umesh, Srinivasan, et al.
Veröffentlicht: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
Dementia classification from spontaneous speech using wrapper-based feature selection
von: Niemelä, Marko, et al.
Veröffentlicht: (2025)
von: Niemelä, Marko, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
von: Thornton, Mike, et al.
Veröffentlicht: (2024) -
Joint decoding method for controllable contextual speech recognition based on Speech LLM
von: Fang, Yangui, et al.
Veröffentlicht: (2025) -
Explainable speech emotion recognition through attentive pooling: insights from attention-based temporal localization
von: Leygue, Tahitoa, et al.
Veröffentlicht: (2025) -
Omni-directional attention mechanism based on Mamba for speech separation
von: Xue, Ke, et al.
Veröffentlicht: (2026) -
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)