Exploring the Power of Pure Attention Mechanisms in Blind Room Parameter Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Chunxi, Jia, Maoshen, Li, Meiran, Bao, Changchun, Jin, Wenyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SS-BRPE: Self-Supervised Blind Room Parameter Estimation Using Attention Mechanisms
von: Wang, Chunxi, et al.
Veröffentlicht: (2024)
von: Wang, Chunxi, et al.
Veröffentlicht: (2024)
DARAS: Dynamic Audio-Room Acoustic Synthesis for Blind Room Impulse Response Estimation
von: Wang, Chunxi, et al.
Veröffentlicht: (2025)
von: Wang, Chunxi, et al.
Veröffentlicht: (2025)
Speech Enhancement with Dual-path Multi-Channel Linear Prediction Filter and Multi-norm Beamforming
von: Qin, Chengyuan, et al.
Veröffentlicht: (2025)
von: Qin, Chengyuan, et al.
Veröffentlicht: (2025)
BERP: A Blind Estimator of Room Parameters for Single-Channel Noisy Speech Signals
von: Wang, Lijun, et al.
Veröffentlicht: (2024)
von: Wang, Lijun, et al.
Veröffentlicht: (2024)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
von: Götz, Philipp, et al.
Veröffentlicht: (2024)
von: Götz, Philipp, et al.
Veröffentlicht: (2024)
Target Speaker Extraction by Directly Exploiting Contextual Information in the Time-Frequency Domain
von: Yang, Xue, et al.
Veröffentlicht: (2024)
von: Yang, Xue, et al.
Veröffentlicht: (2024)
Blind Identification of Binaural Room Impulse Responses from Smart Glasses
von: Deppisch, Thomas, et al.
Veröffentlicht: (2024)
von: Deppisch, Thomas, et al.
Veröffentlicht: (2024)
Blind Localization of Early Room Reflections with Arbitrary Microphone Array
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
Rec-RIR: Monaural Blind Room Impulse Response Identification via DNN-based Reverberant Speech Reconstruction in STFT Domain
von: Wang, Pengyu, et al.
Veröffentlicht: (2025)
von: Wang, Pengyu, et al.
Veröffentlicht: (2025)
Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
Room Impulse Response Generation Conditioned on Acoustic Parameters
von: Arellano, Silvia, et al.
Veröffentlicht: (2025)
von: Arellano, Silvia, et al.
Veröffentlicht: (2025)
Scene-wide Acoustic Parameter Estimation
von: Falcon-Perez, Ricardo, et al.
Veröffentlicht: (2024)
von: Falcon-Perez, Ricardo, et al.
Veröffentlicht: (2024)
AnyRIR: Robust Non-intrusive Room Impulse Response Estimation in the Wild
von: Lee, Kyung Yun, et al.
Veröffentlicht: (2025)
von: Lee, Kyung Yun, et al.
Veröffentlicht: (2025)
Blind Capon Beamformer Based on Independent Component Extraction: Single-Parameter Algorithm,
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2025)
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2025)
Determined Blind Source Separation with Sinkhorn Divergence-based Optimal Allocation of the Source Power
von: Wang, Jianyu, et al.
Veröffentlicht: (2025)
von: Wang, Jianyu, et al.
Veröffentlicht: (2025)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
von: MacWilliam, Kathleen, et al.
Veröffentlicht: (2024)
von: MacWilliam, Kathleen, et al.
Veröffentlicht: (2024)
How Attention Shapes Emotion: A Comparative Study of Attention Mechanisms for Speech Emotion Recognition
von: Casals-Salvador, Marc, et al.
Veröffentlicht: (2026)
von: Casals-Salvador, Marc, et al.
Veröffentlicht: (2026)
Room Impulse Response Estimation through Optimal Mass Transport Barycenters
von: Pallewela, Rumeshika, et al.
Veröffentlicht: (2025)
von: Pallewela, Rumeshika, et al.
Veröffentlicht: (2025)
Multimodal Deep Learning Method for Real-Time Spatial Room Impulse Response Computing
von: Li, Zhiyu, et al.
Veröffentlicht: (2026)
von: Li, Zhiyu, et al.
Veröffentlicht: (2026)
Closed-Form Successive Relative Transfer Function Vector Estimation based on Blind Oblique Projection Incorporating Noise Whitening
von: Gode, Henri, et al.
Veröffentlicht: (2025)
von: Gode, Henri, et al.
Veröffentlicht: (2025)
Room Impulse Response Estimation using Optimal Transport: Simulation-Informed Inference
von: Sundström, David, et al.
Veröffentlicht: (2024)
von: Sundström, David, et al.
Veröffentlicht: (2024)
Speech Emotion Recognition Via CNN-Transformer and Multidimensional Attention Mechanism
von: Tang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tang, Xiaoyu, et al.
Veröffentlicht: (2024)
MEAN-RIR: Multi-Modal Environment-Aware Network for Robust Room Impulse Response Estimation
von: Chen, Jiajian, et al.
Veröffentlicht: (2025)
von: Chen, Jiajian, et al.
Veröffentlicht: (2025)
Deep Room Impulse Response Completion
von: Lin, Jackie, et al.
Veröffentlicht: (2024)
von: Lin, Jackie, et al.
Veröffentlicht: (2024)
MVANet: Multi-Stage Video Attention Network for Sound Event Localization and Detection with Source Distance Estimation
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
von: Hong, Hengyi, et al.
Veröffentlicht: (2024)
Generative Data Augmentation Challenge: Synthesis of Room Acoustics for Speaker Distance Estimation
von: Lin, Jackie, et al.
Veröffentlicht: (2025)
von: Lin, Jackie, et al.
Veröffentlicht: (2025)
Single Channel Blind Dereverberation of Speech Signals
von: Nigam, Dhruv
Veröffentlicht: (2025)
von: Nigam, Dhruv
Veröffentlicht: (2025)
Joint Multi-scale Cross-lingual Speaking Style Transfer with Bidirectional Attention Mechanism for Automatic Dubbing
von: Li, Jingbei, et al.
Veröffentlicht: (2023)
von: Li, Jingbei, et al.
Veröffentlicht: (2023)
Tracking Listener Attention: Gaze-Guided Audio-Visual Speech Enhancement Framework
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
Semi-Blind Channel Estimation and Hybrid Receiver Beamforming in the Tera-Hertz Multi-User Massive MIMO Uplink
von: Garg, Abhisha, et al.
Veröffentlicht: (2026)
von: Garg, Abhisha, et al.
Veröffentlicht: (2026)
Room compensation for loudspeaker reproduction using a supporting source
von: Brooks-Park, James, et al.
Veröffentlicht: (2026)
von: Brooks-Park, James, et al.
Veröffentlicht: (2026)
Blind Spatial Impulse Response Generation from Separate Room- and Scene-Specific Information
von: Lluís, Francesc, et al.
Veröffentlicht: (2024)
von: Lluís, Francesc, et al.
Veröffentlicht: (2024)
Diminishing Domain Mismatch for DNN-Based Acoustic Distance Estimation via Stochastic Room Reverberation Models
von: Gburrek, Tobias, et al.
Veröffentlicht: (2024)
von: Gburrek, Tobias, et al.
Veröffentlicht: (2024)
Pureformer-VC: Non-parallel Voice Conversion with Pure Stylized Transformer Blocks and Triplet Discriminative Training
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
Adapting a Text-to-Audio Model for Room Impulse Response Generation
von: Kim, Kirak, et al.
Veröffentlicht: (2026)
von: Kim, Kirak, et al.
Veröffentlicht: (2026)
Towards Effective and Efficient Non-autoregressive decoders for Conformer and LLM-based ASR using Block-based Attention Mask
von: Wang, Tianzi, et al.
Veröffentlicht: (2025)
von: Wang, Tianzi, et al.
Veröffentlicht: (2025)
Towards Blind Data Cleaning: A Case Study in Music Source Separation
von: Gui, Azalea, et al.
Veröffentlicht: (2025)
von: Gui, Azalea, et al.
Veröffentlicht: (2025)
Determined Multichannel Blind Source Separation with Clustered Source Model
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
von: Wang, Jianyu, et al.
Veröffentlicht: (2024)
Blind Estimation of Sub-band Acoustic Parameters from Ambisonics Recordings using Spectro-Spatial Covariance Features
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
Learning Filters in Feedback Delay Networks from Noisy Room Impulse Responses
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2025)
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SS-BRPE: Self-Supervised Blind Room Parameter Estimation Using Attention Mechanisms
von: Wang, Chunxi, et al.
Veröffentlicht: (2024) -
DARAS: Dynamic Audio-Room Acoustic Synthesis for Blind Room Impulse Response Estimation
von: Wang, Chunxi, et al.
Veröffentlicht: (2025) -
Speech Enhancement with Dual-path Multi-Channel Linear Prediction Filter and Multi-norm Beamforming
von: Qin, Chengyuan, et al.
Veröffentlicht: (2025) -
BERP: A Blind Estimator of Room Parameters for Single-Channel Noisy Speech Signals
von: Wang, Lijun, et al.
Veröffentlicht: (2024) -
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
von: Götz, Philipp, et al.
Veröffentlicht: (2024)