Deep Learning-Based Approach for Identification and Compensation of Nonlinear Distortions in Parametric Array Loudspeakers
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Mengtong, Zhuang, Tao, Chen, Kai, Zhong, Jia-Xin, Lu, Jing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Generating Localized Audible Zones Using a Single-Channel Parametric Loudspeaker
por: Zhuang, Tao, et al.
Publicado: (2025)
por: Zhuang, Tao, et al.
Publicado: (2025)
A k-space approach to modeling multi-channel parametric array loudspeaker systems
por: Zhuang, Tao, et al.
Publicado: (2025)
por: Zhuang, Tao, et al.
Publicado: (2025)
Optimized Loudspeaker Panning for Adaptive Sound-Field Correction and Non-stationary Listening Areas
por: Luo, Yuancheng
Publicado: (2025)
por: Luo, Yuancheng
Publicado: (2025)
Constant Directivity Loudspeaker Beamforming
por: Luo, Yuancheng
Publicado: (2024)
por: Luo, Yuancheng
Publicado: (2024)
Loudspeaker Beamforming to Enhance Speech Recognition Performance of Voice Driven Applications
por: de Groot, Dimme, et al.
Publicado: (2025)
por: de Groot, Dimme, et al.
Publicado: (2025)
Inverse Nonlinearity Compensation of Hyperelastic Deformation in Dielectric Elastomer for Acoustic Actuation
por: Lee, Jin Woo, et al.
Publicado: (2024)
por: Lee, Jin Woo, et al.
Publicado: (2024)
Direction Estimation of Sound Sources Using Microphone Arrays and Signal Strength
por: Pour, Mahdi Ali, et al.
Publicado: (2025)
por: Pour, Mahdi Ali, et al.
Publicado: (2025)
SNR-Progressive Model with Harmonic Compensation for Low-SNR Speech Enhancement
por: Hou, Zhongshu, et al.
Publicado: (2024)
por: Hou, Zhongshu, et al.
Publicado: (2024)
Reduction of Nonlinear Distortion in Condenser Microphones Using a Simple Post-Processing Technique
por: Honzík, Petr, et al.
Publicado: (2024)
por: Honzík, Petr, et al.
Publicado: (2024)
Deep Learning Based Stage-wise Two-dimensional Speaker Localization with Large Ad-hoc Microphone Arrays
por: Liu, Shupei, et al.
Publicado: (2022)
por: Liu, Shupei, et al.
Publicado: (2022)
UniArray: Unified Spectral-Spatial Modeling for Array-Geometry-Agnostic Speech Separation
por: Chen, Weiguang, et al.
Publicado: (2025)
por: Chen, Weiguang, et al.
Publicado: (2025)
Generalized Fake Audio Detection via Deep Stable Learning
por: Wang, Zhiyong, et al.
Publicado: (2024)
por: Wang, Zhiyong, et al.
Publicado: (2024)
Synthetic Singers: A Review of Deep-Learning-based Singing Voice Synthesis Approaches
por: Pan, Changhao, et al.
Publicado: (2026)
por: Pan, Changhao, et al.
Publicado: (2026)
VM-UNSSOR: Unsupervised Neural Speech Separation Enhanced by Higher-SNR Virtual Microphone Arrays
por: He, Shulin, et al.
Publicado: (2025)
por: He, Shulin, et al.
Publicado: (2025)
Broadband MEMS Microphone Arrays with Reduced Aperture Through 3D-Printed Waveguides
por: Laurijssen, Dennis, et al.
Publicado: (2024)
por: Laurijssen, Dennis, et al.
Publicado: (2024)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
Automatic Live Music Song Identification Using Multi-level Deep Sequence Similarity Learning
por: Hakala, Aapo, et al.
Publicado: (2025)
por: Hakala, Aapo, et al.
Publicado: (2025)
End-to-End Multi-Task Learning for Adjustable Joint Noise Reduction and Hearing Loss Compensation
por: Gonzalez, Philippe, et al.
Publicado: (2026)
por: Gonzalez, Philippe, et al.
Publicado: (2026)
Frequency Tracking Features for Data-Efficient Deep Siren Identification
por: Damiano, Stefano, et al.
Publicado: (2024)
por: Damiano, Stefano, et al.
Publicado: (2024)
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
por: Wen, Wen, et al.
Publicado: (2024)
por: Wen, Wen, et al.
Publicado: (2024)
Array Geometry-Robust Attention-Based Neural Beamformer for Moving Speakers
por: Tammen, Marvin, et al.
Publicado: (2024)
por: Tammen, Marvin, et al.
Publicado: (2024)
Masked Audio Modeling with CLAP and Multi-Objective Learning
por: Xin, Yifei, et al.
Publicado: (2024)
por: Xin, Yifei, et al.
Publicado: (2024)
Deep Learning for Personalized Binaural Audio Reproduction
por: Lu, Xikun, et al.
Publicado: (2025)
por: Lu, Xikun, et al.
Publicado: (2025)
Distortion Recovery: A Two-Stage Method for Guitar Effect Removal
por: Lee, Ying-Shuo, et al.
Publicado: (2024)
por: Lee, Ying-Shuo, et al.
Publicado: (2024)
Magnetoencephalography (MEG) Based Non-Invasive Chinese Speech Decoding
por: Jia, Zhihong, et al.
Publicado: (2025)
por: Jia, Zhihong, et al.
Publicado: (2025)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Rethinking Processing Distortions: Disentangling the Impact of Speech Enhancement Errors on Speech Recognition Performance
por: Ochiai, Tsubasa, et al.
Publicado: (2024)
por: Ochiai, Tsubasa, et al.
Publicado: (2024)
ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation
por: Fu, Ruibo, et al.
Publicado: (2024)
por: Fu, Ruibo, et al.
Publicado: (2024)
Uncertainty Quantification in Machine Learning for Joint Speaker Diarization and Identification
por: McKnight, Simon W., et al.
Publicado: (2023)
por: McKnight, Simon W., et al.
Publicado: (2023)
Comparative Analysis Of Discriminative Deep Learning-Based Noise Reduction Methods In Low SNR Scenarios
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
The Impact of Frequency Bands on Acoustic Anomaly Detection of Machines using Deep Learning Based Model
por: Nguyen, Tin, et al.
Publicado: (2024)
por: Nguyen, Tin, et al.
Publicado: (2024)
WMCodec: End-to-End Neural Speech Codec with Deep Watermarking for Authenticity Verification
por: Zhou, Junzuo, et al.
Publicado: (2024)
por: Zhou, Junzuo, et al.
Publicado: (2024)
Assisted RTF-Vector-Based Binaural Direction of Arrival Estimation Exploiting a Calibrated External Microphone Array
por: Fejgin, Daniel, et al.
Publicado: (2022)
por: Fejgin, Daniel, et al.
Publicado: (2022)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
por: Yang, Bing, et al.
Publicado: (2024)
por: Yang, Bing, et al.
Publicado: (2024)
SE/BN Adapter: Parametric Efficient Domain Adaptation for Speaker Recognition
por: Wang, Tianhao, et al.
Publicado: (2024)
por: Wang, Tianhao, et al.
Publicado: (2024)
ctPuLSE: Close-Talk, and Pseudo-Label Based Far-Field, Speech Enhancement
por: Wang, Zhong-Qiu
Publicado: (2024)
por: Wang, Zhong-Qiu
Publicado: (2024)
Pitch Contour Exploration Across Audio Domains: A Vision-Based Transfer Learning Approach
por: Abeßer, Jakob, et al.
Publicado: (2025)
por: Abeßer, Jakob, et al.
Publicado: (2025)
SoCov: Semi-Orthogonal Parametric Pooling of Covariance Matrix for Speaker Recognition
por: Li, Rongjin, et al.
Publicado: (2025)
por: Li, Rongjin, et al.
Publicado: (2025)
EffectiveASR: A Single-Step Non-Autoregressive Mandarin Speech Recognition Architecture with High Accuracy and Inference Speed
por: Zhuang, Ziyang, et al.
Publicado: (2024)
por: Zhuang, Ziyang, et al.
Publicado: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Ejemplares similares
-
Generating Localized Audible Zones Using a Single-Channel Parametric Loudspeaker
por: Zhuang, Tao, et al.
Publicado: (2025) -
A k-space approach to modeling multi-channel parametric array loudspeaker systems
por: Zhuang, Tao, et al.
Publicado: (2025) -
Optimized Loudspeaker Panning for Adaptive Sound-Field Correction and Non-stationary Listening Areas
por: Luo, Yuancheng
Publicado: (2025) -
Constant Directivity Loudspeaker Beamforming
por: Luo, Yuancheng
Publicado: (2024) -
Loudspeaker Beamforming to Enhance Speech Recognition Performance of Voice Driven Applications
por: de Groot, Dimme, et al.
Publicado: (2025)