Neural Ambisonic Encoding For Multi-Speaker Scenarios Using A Circular Microphone Array
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiao, Yue, Kothapally, Vinay, Yu, Meng, Yu, Dong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ambisonics Encoding For Arbitrary Microphone Arrays Incorporating Residual Channels For Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024)
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
Beyond Omnidirectional: Neural Ambisonics Encoding for Arbitrary Microphone Directivity Patterns using Cross-Attention
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2026)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2026)
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
von: Wen, Wen, et al.
Veröffentlicht: (2024)
von: Wen, Wen, et al.
Veröffentlicht: (2024)
SpatialCodec: Neural Spatial Speech Coding
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)
AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
Neural Directional Filtering Using a Compact Microphone Array
von: Huang, Weilong, et al.
Veröffentlicht: (2025)
von: Huang, Weilong, et al.
Veröffentlicht: (2025)
Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment
von: Shao, Yiwen, et al.
Veröffentlicht: (2024)
von: Shao, Yiwen, et al.
Veröffentlicht: (2024)
GAN-Based Multi-Microphone Spatial Target Speaker Extraction
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2025)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2025)
SpatialEmb: Extract and Encode Spatial Information for 1-Stage Multi-channel Multi-speaker ASR on Arbitrary Microphone Arrays
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
von: Shao, Yiwen, et al.
Veröffentlicht: (2026)
Microphone Occlusion Mitigation for Own-Voice Enhancement in Head-Worn Microphone Arrays Using Switching-Adaptive Beamforming
von: Middelberg, Wiebke, et al.
Veröffentlicht: (2025)
von: Middelberg, Wiebke, et al.
Veröffentlicht: (2025)
Target Speaker Selection for Neural Network Beamforming in Multi-Speaker Scenarios
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
What Does the Speaker Embedding Encode?
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Ambisonics Encoder for Wearable Array with Improved Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
Advances in Microphone Array Processing and Multichannel Speech Enhancement
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
Ambisonics Super-Resolution Using A Waveform-Domain Neural Network
von: Nawfal, Ismael, et al.
Veröffentlicht: (2025)
von: Nawfal, Ismael, et al.
Veröffentlicht: (2025)
Deep Learning Based Stage-wise Two-dimensional Speaker Localization with Large Ad-hoc Microphone Arrays
von: Liu, Shupei, et al.
Veröffentlicht: (2022)
von: Liu, Shupei, et al.
Veröffentlicht: (2022)
Residual Learning for Neural Ambisonics Encoders
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
von: Deppisch, Thomas, et al.
Veröffentlicht: (2026)
SonicBoom: Contact Localization Using Array of Microphones
von: Lee, Moonyoung, et al.
Veröffentlicht: (2024)
von: Lee, Moonyoung, et al.
Veröffentlicht: (2024)
Your Microphone Array Retains Your Identity: A Robust Voice Liveness Detection System for Smart Speakers
von: Meng, Yan, et al.
Veröffentlicht: (2025)
von: Meng, Yan, et al.
Veröffentlicht: (2025)
Linearly Constrained Deep Beamformer for Multi-Speaker Scenarios
von: Zaidel, Ilai, et al.
Veröffentlicht: (2026)
von: Zaidel, Ilai, et al.
Veröffentlicht: (2026)
Ambisonizer: Neural Upmixing as Spherical Harmonics Generation
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
Hierarchical Sparse Sound Field Reconstruction with Spherical and Linear Microphone Arrays
von: Xu, Shunxi, et al.
Veröffentlicht: (2025)
von: Xu, Shunxi, et al.
Veröffentlicht: (2025)
Neural Directional Filtering: Far-Field Directivity Control With a Small Microphone Array
von: Wechsler, Julian, et al.
Veröffentlicht: (2024)
von: Wechsler, Julian, et al.
Veröffentlicht: (2024)
Neural Ambisonics encoding for compact irregular microphone arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
VM-UNSSOR: Unsupervised Neural Speech Separation Enhanced by Higher-SNR Virtual Microphone Arrays
von: He, Shulin, et al.
Veröffentlicht: (2025)
von: He, Shulin, et al.
Veröffentlicht: (2025)
RIR-SF: Room Impulse Response Based Spatial Feature for Target Speech Recognition in Multi-Channel Multi-Speaker Scenarios
von: Shao, Yiwen, et al.
Veröffentlicht: (2023)
von: Shao, Yiwen, et al.
Veröffentlicht: (2023)
Single-Microphone Speaker Separation and Voice Activity Detection in Noisy and Reverberant Environments
von: Opochinsky, Renana, et al.
Veröffentlicht: (2024)
von: Opochinsky, Renana, et al.
Veröffentlicht: (2024)
Design and Analysis of Binaural Signal Matching with Arbitrary Microphone Arrays and Listener Head Rotations
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
von: Madmoni, Lior, et al.
Veröffentlicht: (2024)
Applying Automatic Differentiation to Optimize Differential Microphone Array Designs
von: Galougah, Siminfar Samakoush, et al.
Veröffentlicht: (2024)
von: Galougah, Siminfar Samakoush, et al.
Veröffentlicht: (2024)
Asynchronous Microphone Array Calibration using Hybrid TDOA Information
von: Zhang, Chengjie, et al.
Veröffentlicht: (2024)
von: Zhang, Chengjie, et al.
Veröffentlicht: (2024)
Blind Localization of Early Room Reflections with Arbitrary Microphone Array
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
von: Hadadi, Yogev, et al.
Veröffentlicht: (2024)
Evaluation of Spherical Wavelet Framework in Comparsion with Ambisonics
von: Ekmen, Ş., et al.
Veröffentlicht: (2025)
von: Ekmen, Ş., et al.
Veröffentlicht: (2025)
Introduction to Ambisonics, Part 1: The Part With No Math
von: Ahrens, Jens
Veröffentlicht: (2025)
von: Ahrens, Jens
Veröffentlicht: (2025)
Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Fast Multichannel NMF with Block-Diagonal Spatial Covariance Matrices for Efficient Blind Source Separation Using Distributed Microphone Arrays
von: Nishikori, Hirotaka, et al.
Veröffentlicht: (2026)
von: Nishikori, Hirotaka, et al.
Veröffentlicht: (2026)
Direction of Arrival Estimation Using Microphone Array Processing for Moving Humanoid Robots
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
A Unified SVD-Modal Solution for Sparse Sound Field Reconstruction with Hybrid Spherical-Linear Microphone Arrays
von: Xu, Shunxi, et al.
Veröffentlicht: (2026)
von: Xu, Shunxi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Ambisonics Encoding For Arbitrary Microphone Arrays Incorporating Residual Channels For Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2024) -
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025) -
Beyond Omnidirectional: Neural Ambisonics Encoding for Arbitrary Microphone Directivity Patterns using Cross-Attention
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2026) -
Neural Directed Speech Enhancement with Dual Microphone Array in High Noise Scenario
von: Wen, Wen, et al.
Veröffentlicht: (2024) -
SpatialCodec: Neural Spatial Speech Coding
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2023)