Beyond Omnidirectional: Neural Ambisonics Encoding for Arbitrary Microphone Directivity Patterns using Cross-Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Heikkinen, Mikko, Politis, Archontis, Drossos, Konstantinos, Virtanen, Tuomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
Neural Ambisonics encoding for compact irregular microphone arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
Moving Speaker Separation via Parallel Spectral-Spatial Processing
von: Wang, Yuzhu, et al.
Veröffentlicht: (2026)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2026)
Multi-Utterance Speech Separation and Association Trained on Short Segments
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Attractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Automatic Contextual Audio Denoising
von: Luong, Diep, et al.
Veröffentlicht: (2026)
von: Luong, Diep, et al.
Veröffentlicht: (2026)
Knowledge Distillation for Speech Denoising by Latent Representation Alignment with Cosine Distance
von: Luong, Diep, et al.
Veröffentlicht: (2025)
von: Luong, Diep, et al.
Veröffentlicht: (2025)
Inter-Speaker Relative Cues for Two-Stage Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2026)
von: Dai, Wang, et al.
Veröffentlicht: (2026)
Inter-Speaker Relative Cues for Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2025)
von: Dai, Wang, et al.
Veröffentlicht: (2025)
Gaunt coefficients for complex and real spherical harmonics with applications to spherical array processing and Ambisonics
von: Politis, Archontis
Veröffentlicht: (2024)
von: Politis, Archontis
Veröffentlicht: (2024)
Reference Channel Selection by Multi-Channel Masking for End-to-End Multi-Channel Speech Enhancement
von: Dai, Wang, et al.
Veröffentlicht: (2024)
von: Dai, Wang, et al.
Veröffentlicht: (2024)
Lightweight DNN for Full-Band Speech Denoising on Mobile Devices: Exploiting Long and Short Temporal Patterns
von: Drossos, Konstantinos, et al.
Veröffentlicht: (2025)
von: Drossos, Konstantinos, et al.
Veröffentlicht: (2025)
Dependence on Early and Late Reverberation of Single-Channel Speaker Distance Estimation
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Ambisonics Encoding For Arbitrary Microphone Arrays Incorporating Residual Channels For Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
Discriminating real and synthetic super-resolved audio samples using embedding-based classifiers
von: Silaev, Mikhail, et al.
Veröffentlicht: (2026)
von: Silaev, Mikhail, et al.
Veröffentlicht: (2026)
Perceptually-motivated Spatial Audio Codec for Higher-Order Ambisonics Compression
von: Hold, Christoph, et al.
Veröffentlicht: (2024)
von: Hold, Christoph, et al.
Veröffentlicht: (2024)
Speaker Distance Estimation in Enclosures from Single-Channel Audio
von: Neri, Michael, et al.
Veröffentlicht: (2024)
von: Neri, Michael, et al.
Veröffentlicht: (2024)
Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
Neural Ambisonic Encoding For Multi-Speaker Scenarios Using A Circular Microphone Array
von: Qiao, Yue, et al.
Veröffentlicht: (2024)
von: Qiao, Yue, et al.
Veröffentlicht: (2024)
Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
Towards Spatial Audio Understanding via Question Answering
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
Representation Learning for Audio Privacy Preservation using Source Separation and Robust Adversarial Learning
von: Luong, Diep, et al.
Veröffentlicht: (2023)
von: Luong, Diep, et al.
Veröffentlicht: (2023)
Adversarial Representation Learning for Robust Privacy Preservation in Audio
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
Class-Incremental Learning for Sound Event Localization and Detection
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
von: Pandey, Ruchi, et al.
Veröffentlicht: (2024)
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
SynthSOD: Developing an Heterogeneous Dataset for Orchestra Music Source Separation
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2024)
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2024)
Noise-to-mask Ratio Loss for Deep Neural Network based Audio Watermarking
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
Neural Directional Filtering Using a Compact Microphone Array
von: Huang, Weilong, et al.
Veröffentlicht: (2025)
von: Huang, Weilong, et al.
Veröffentlicht: (2025)
Sound Event Detection and Localization with Distance Estimation
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
von: Krause, Daniel Aleksander, et al.
Veröffentlicht: (2024)
Multi-Channel Replay Speech Detection using Acoustic Maps
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
Residual Learning for Neural Ambisonics Encoders
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
Blind Localization of Early Room Reflections with Arbitrary Microphone Array
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
Design and Analysis of Binaural Signal Matching with Arbitrary Microphone Arrays and Listener Head Rotations
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
Ambisonizer: Neural Upmixing as Spherical Harmonics Generation
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
Automatic Live Music Song Identification Using Multi-level Deep Sequence Similarity Learning
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
Integrating Continuous and Binary Relevances in Audio-Text Relevance Learning
von: Xie, Huang, et al.
Veröffentlicht: (2024)
von: Xie, Huang, et al.
Veröffentlicht: (2024)
Text-based Audio Retrieval by Learning from Similarities between Audio Captions
von: Xie, Huang, et al.
Veröffentlicht: (2024)
von: Xie, Huang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025) -
Neural Ambisonics encoding for compact irregular microphone arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024) -
Moving Speaker Separation via Parallel Spectral-Spatial Processing
von: Wang, Yuzhu, et al.
Veröffentlicht: (2026) -
Multi-Utterance Speech Separation and Association Trained on Short Segments
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025) -
Attractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)