Scattering Transform for Auditory Attention Decoding
Fuente:
arXiv
Salvato in:
| Autori principali: | Pallenberg, René, Katzberg, Fabrice, Mertins, Alfred, Maass, Marco |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding
di: Zhang, Ziyang, et al.
Pubblicazione: (2024)
di: Zhang, Ziyang, et al.
Pubblicazione: (2024)
Auditory Attention Decoding from Ear-EEG Signals: A Dataset with Dynamic Attention Switching and Rigorous Cross-Validation
di: Zhang, Yuanming, et al.
Pubblicazione: (2025)
di: Zhang, Yuanming, et al.
Pubblicazione: (2025)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
di: Zhu, Haolin, et al.
Pubblicazione: (2024)
di: Zhu, Haolin, et al.
Pubblicazione: (2024)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024)
Single-word Auditory Attention Decoding Using Deep Learning Model
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024)
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
di: Geirnaert, Simon, et al.
Pubblicazione: (2025)
di: Geirnaert, Simon, et al.
Pubblicazione: (2025)
Toward Fully-End-to-End Listened Speech Decoding from EEG Signals
di: Lee, Jihwan, et al.
Pubblicazione: (2024)
di: Lee, Jihwan, et al.
Pubblicazione: (2024)
EEG-Based Speech Decoding: A Novel Approach Using Multi-Kernel Ensemble Diffusion Models
di: Kim, Soowon, et al.
Pubblicazione: (2024)
di: Kim, Soowon, et al.
Pubblicazione: (2024)
A Domain-Knowledge-Inspired Music Embedding Space and a Novel Attention Mechanism for Symbolic Music Modeling
di: Guo, Z., et al.
Pubblicazione: (2022)
di: Guo, Z., et al.
Pubblicazione: (2022)
Privacy-Preserving End-to-End Full-Duplex Speech Dialogue Models
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
StreamVoiceAnon+: Emotion-Preserving Streaming Speaker Anonymization via Frame-Level Acoustic Distillation
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
Explainable AI in Speaker Recognition -- Making Latent Representations Understandable
di: Xu, Yanze, et al.
Pubblicazione: (2026)
di: Xu, Yanze, et al.
Pubblicazione: (2026)
Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment
di: Cai, Huanchen, et al.
Pubblicazione: (2026)
di: Cai, Huanchen, et al.
Pubblicazione: (2026)
Auditory Attention Decoding without Spatial Information: A Diotic EEG Study
di: Yoshino, Masahiro, et al.
Pubblicazione: (2026)
di: Yoshino, Masahiro, et al.
Pubblicazione: (2026)
Transferable Adversarial Attacks against ASR
di: Gao, Xiaoxue, et al.
Pubblicazione: (2024)
di: Gao, Xiaoxue, et al.
Pubblicazione: (2024)
Crowdsourced Multilingual Speech Intelligibility Testing
di: Lechler, Laura, et al.
Pubblicazione: (2024)
di: Lechler, Laura, et al.
Pubblicazione: (2024)
Mitigating Intra-Speaker Variability in Diarization with Style-Controllable Speech Augmentation
di: Kim, Miseul, et al.
Pubblicazione: (2025)
di: Kim, Miseul, et al.
Pubblicazione: (2025)
An Investigation of Noise Robustness for Flow-Matching-Based Zero-Shot TTS
di: Wang, Xiaofei, et al.
Pubblicazione: (2024)
di: Wang, Xiaofei, et al.
Pubblicazione: (2024)
Laugh Now Cry Later: Controlling Time-Varying Emotional States of Flow-Matching-Based Zero-Shot Text-to-Speech
di: Wu, Haibin, et al.
Pubblicazione: (2024)
di: Wu, Haibin, et al.
Pubblicazione: (2024)
FreGrad: Lightweight and Fast Frequency-aware Diffusion Vocoder
di: Nguyen, Tan Dat, et al.
Pubblicazione: (2024)
di: Nguyen, Tan Dat, et al.
Pubblicazione: (2024)
Text-Aware Adapter for Few-Shot Keyword Spotting
di: Jung, Youngmoon, et al.
Pubblicazione: (2024)
di: Jung, Youngmoon, et al.
Pubblicazione: (2024)
Neural Spectral Band Generation for Audio Coding
di: Choi, Woongjib, et al.
Pubblicazione: (2025)
di: Choi, Woongjib, et al.
Pubblicazione: (2025)
Relational Proxy Loss for Audio-Text based Keyword Spotting
di: Jung, Youngmoon, et al.
Pubblicazione: (2024)
di: Jung, Youngmoon, et al.
Pubblicazione: (2024)
Enhancing dysarthria speech feature representation with empirical mode decomposition and Walsh-Hadamard transform
di: Zhu, Ting, et al.
Pubblicazione: (2023)
di: Zhu, Ting, et al.
Pubblicazione: (2023)
CSL-L2M: Controllable Song-Level Lyric-to-Melody Generation Based on Conditional Transformer with Fine-Grained Lyric and Musical Controls
di: Chai, Li, et al.
Pubblicazione: (2024)
di: Chai, Li, et al.
Pubblicazione: (2024)
Enhancing Listened Speech Decoding from EEG via Parallel Phoneme Sequence Prediction
di: Lee, Jihwan, et al.
Pubblicazione: (2025)
di: Lee, Jihwan, et al.
Pubblicazione: (2025)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
di: Liao, Yuan, et al.
Pubblicazione: (2025)
di: Liao, Yuan, et al.
Pubblicazione: (2025)
Decoding Selective Auditory Attention to Musical Elements in Ecologically Valid Music Listening
di: Akama, Taketo, et al.
Pubblicazione: (2025)
di: Akama, Taketo, et al.
Pubblicazione: (2025)
A Knowledge-Driven Approach to Music Segmentation, Music Source Separation and Cinematic Audio Source Separation
di: Ho, Chun-wei, et al.
Pubblicazione: (2026)
di: Ho, Chun-wei, et al.
Pubblicazione: (2026)
AI-Generated Music Detection in Broadcast Monitoring
di: López-Ayala, David, et al.
Pubblicazione: (2026)
di: López-Ayala, David, et al.
Pubblicazione: (2026)
Speech Enhancement Based on Drifting Models
di: Xu, Liang, et al.
Pubblicazione: (2026)
di: Xu, Liang, et al.
Pubblicazione: (2026)
Discriminating real and synthetic super-resolved audio samples using embedding-based classifiers
di: Silaev, Mikhail, et al.
Pubblicazione: (2026)
di: Silaev, Mikhail, et al.
Pubblicazione: (2026)
Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation
di: Ratnarajah, Anton, et al.
Pubblicazione: (2026)
di: Ratnarajah, Anton, et al.
Pubblicazione: (2026)
Dependence on Early and Late Reverberation of Single-Channel Speaker Distance Estimation
di: Neri, Michael, et al.
Pubblicazione: (2026)
di: Neri, Michael, et al.
Pubblicazione: (2026)
Robust Generative Audio Quality Assessment: Disentangling Quality from Spurious Correlations
di: Huang, Kuan-Tang, et al.
Pubblicazione: (2026)
di: Huang, Kuan-Tang, et al.
Pubblicazione: (2026)
Compressing Quaternion Convolutional Neural Networks for Audio Classification
di: Singh, Arshdeep, et al.
Pubblicazione: (2025)
di: Singh, Arshdeep, et al.
Pubblicazione: (2025)
Construction and Evaluation of Mandarin Multimodal Emotional Speech Database
di: Ting, Zhu, et al.
Pubblicazione: (2024)
di: Ting, Zhu, et al.
Pubblicazione: (2024)
On the Parameter Estimation of Sinusoidal Models for Speech and Audio Signals
di: Kafentzis, George P.
Pubblicazione: (2024)
di: Kafentzis, George P.
Pubblicazione: (2024)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
di: Kim, Minje, et al.
Pubblicazione: (2024)
di: Kim, Minje, et al.
Pubblicazione: (2024)
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
di: Lee, Jin Woo, et al.
Pubblicazione: (2024)
di: Lee, Jin Woo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding
di: Zhang, Ziyang, et al.
Pubblicazione: (2024) -
Auditory Attention Decoding from Ear-EEG Signals: A Dataset with Dynamic Attention Switching and Rigorous Cross-Validation
di: Zhang, Yuanming, et al.
Pubblicazione: (2025) -
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
di: Zhu, Haolin, et al.
Pubblicazione: (2024) -
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024) -
Single-word Auditory Attention Decoding Using Deep Learning Model
di: Nguyen, Nhan Duc Thanh, et al.
Pubblicazione: (2024)