Guardado en:
| Autores principales: | Elbir, Ahmet M., Demir, Özlem Tuğfe, Mishra, Kumar Vijay, Chatzinotas, Symeon, Haardt, Martin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2408.11434 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
AI-Driven Cardiorespiratory Signal Processing: Separation, Clustering, and Anomaly Detection
por: Torabi, Yasaman
Publicado: (2026)
por: Torabi, Yasaman
Publicado: (2026)
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
por: Abbott, Leigh, et al.
Publicado: (2024)
por: Abbott, Leigh, et al.
Publicado: (2024)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
por: Meng, Hanyu, et al.
Publicado: (2025)
por: Meng, Hanyu, et al.
Publicado: (2025)
Phase-Based Signal Representations for Scattering
por: Haider, Daniel, et al.
Publicado: (2022)
por: Haider, Daniel, et al.
Publicado: (2022)
Blind Source Separation of Radar Signals in Time Domain Using Deep Learning
por: Hinderer, Sven
Publicado: (2025)
por: Hinderer, Sven
Publicado: (2025)
Comparison of Classification Algorithms for COVID19 Detection using Cough Acoustic Signals
por: Erdoğan, Yunus Emre, et al.
Publicado: (2022)
por: Erdoğan, Yunus Emre, et al.
Publicado: (2022)
Physics-Informed Direction-Aware Neural Acoustic Fields
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
por: Liao, Yuan, et al.
Publicado: (2025)
por: Liao, Yuan, et al.
Publicado: (2025)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
por: Brümann, Klaus, et al.
Publicado: (2024)
por: Brümann, Klaus, et al.
Publicado: (2024)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
por: Masuyama, Yoshiki, et al.
Publicado: (2026)
por: Masuyama, Yoshiki, et al.
Publicado: (2026)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
por: Carson, Alistair, et al.
Publicado: (2024)
por: Carson, Alistair, et al.
Publicado: (2024)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
Cross-Talk Reduction
por: Wang, Zhong-Qiu, et al.
Publicado: (2024)
por: Wang, Zhong-Qiu, et al.
Publicado: (2024)
Hybrid Deep Learning and Signal Processing for Arabic Dialect Recognition in Low-Resource Settings
por: Al-Shwayyat, Ghazal, et al.
Publicado: (2025)
por: Al-Shwayyat, Ghazal, et al.
Publicado: (2025)
DeWinder: Single-Channel Wind Noise Reduction using Ultrasound Sensing
por: Yuan, Kuang, et al.
Publicado: (2024)
por: Yuan, Kuang, et al.
Publicado: (2024)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription
por: Zang, Yongyi, et al.
Publicado: (2023)
por: Zang, Yongyi, et al.
Publicado: (2023)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
por: Wilroth, Johanna, et al.
Publicado: (2026)
por: Wilroth, Johanna, et al.
Publicado: (2026)
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Ternary Spike-based Neuromorphic Signal Processing System
por: Wang, Shuai, et al.
Publicado: (2024)
por: Wang, Shuai, et al.
Publicado: (2024)
Time-domain sound field estimation using kernel ridge regression
por: Brunnström, Jesper, et al.
Publicado: (2025)
por: Brunnström, Jesper, et al.
Publicado: (2025)
Point Processes and spatial statistics in time-frequency analysis
por: Pascal, Barbara, et al.
Publicado: (2024)
por: Pascal, Barbara, et al.
Publicado: (2024)
Lessons Learned from the URGENT 2024 Speech Enhancement Challenge
por: Zhang, Wangyou, et al.
Publicado: (2025)
por: Zhang, Wangyou, et al.
Publicado: (2025)
FoVNet: Configurable Field-of-View Speech Enhancement with Low Computation and Distortion for Smart Glasses
por: Xu, Zhongweiyang, et al.
Publicado: (2024)
por: Xu, Zhongweiyang, et al.
Publicado: (2024)
Task and Perception-aware Distributed Source Coding for Correlated Speech under Bandwidth-constrained Channels
por: Bhattacharya, Sagnik, et al.
Publicado: (2025)
por: Bhattacharya, Sagnik, et al.
Publicado: (2025)
Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
por: Sechaud, Victor, et al.
Publicado: (2024)
por: Sechaud, Victor, et al.
Publicado: (2024)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
por: Gayer, Yhonatan, et al.
Publicado: (2025)
por: Gayer, Yhonatan, et al.
Publicado: (2025)
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
por: Picard, Emilio, et al.
Publicado: (2026)
por: Picard, Emilio, et al.
Publicado: (2026)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
por: Waheed, Muhammad Fasih, et al.
Publicado: (2026)
por: Waheed, Muhammad Fasih, et al.
Publicado: (2026)
Singing Voice Graph Modeling for SingFake Detection
por: Chen, Xuanjun, et al.
Publicado: (2024)
por: Chen, Xuanjun, et al.
Publicado: (2024)
AutoMashup: Automatic Music Mashups Creation
por: Delabaere, Marine, et al.
Publicado: (2025)
por: Delabaere, Marine, et al.
Publicado: (2025)
Beamforming in the Reproducing Kernel Domain Based on Spatial Differentiation
por: Iwami, Takahiro, et al.
Publicado: (2025)
por: Iwami, Takahiro, et al.
Publicado: (2025)
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models
por: Yang, Ningyuan, et al.
Publicado: (2026)
por: Yang, Ningyuan, et al.
Publicado: (2026)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
por: Yi, Jayeon, et al.
Publicado: (2024)
por: Yi, Jayeon, et al.
Publicado: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
por: Qi, Tianhua, et al.
Publicado: (2024)
por: Qi, Tianhua, et al.
Publicado: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
por: Joysingh, S. Johanan, et al.
Publicado: (2024)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
por: Delgado, Pablo M., et al.
Publicado: (2024)
por: Delgado, Pablo M., et al.
Publicado: (2024)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
por: Iatariene, Taous, et al.
Publicado: (2025)
por: Iatariene, Taous, et al.
Publicado: (2025)
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
por: Cheema, Ahsan J., et al.
Publicado: (2025)
por: Cheema, Ahsan J., et al.
Publicado: (2025)
Ejemplares similares
-
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025) -
AI-Driven Cardiorespiratory Signal Processing: Separation, Clustering, and Anomaly Detection
por: Torabi, Yasaman
Publicado: (2026) -
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
por: Abbott, Leigh, et al.
Publicado: (2024) -
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
por: Meng, Hanyu, et al.
Publicado: (2025) -
Phase-Based Signal Representations for Scattering
por: Haider, Daniel, et al.
Publicado: (2022)