Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
Fuente:
arXiv
Guardado en:
| Autores principales: | Masuyama, Yoshiki, Germain, Francois G., Wichern, Gordon, Hori, Chiori, Roux, Jonathan Le |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Physics-Informed Direction-Aware Neural Acoustic Fields
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
Direction-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses
por: Ick, Christopher, et al.
Publicado: (2025)
por: Ick, Christopher, et al.
Publicado: (2025)
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
FasTUSS: Faster Task-Aware Unified Source Separation
por: Paissan, Francesco, et al.
Publicado: (2025)
por: Paissan, Francesco, et al.
Publicado: (2025)
Exploring Disentangled Neural Speech Codecs from Self-Supervised Representations
por: Aihara, Ryo, et al.
Publicado: (2025)
por: Aihara, Ryo, et al.
Publicado: (2025)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
SUNAC: Source-aware Unified Neural Audio Codec
por: Aihara, Ryo, et al.
Publicado: (2025)
por: Aihara, Ryo, et al.
Publicado: (2025)
Data Augmentation Using Neural Acoustic Fields With Retrieval-Augmented Pre-training
por: Ick, Christopher, et al.
Publicado: (2025)
por: Ick, Christopher, et al.
Publicado: (2025)
Sound Event Bounding Boxes
por: Ebbers, Janek, et al.
Publicado: (2024)
por: Ebbers, Janek, et al.
Publicado: (2024)
Enhanced Reverberation as Supervision for Unsupervised Speech Separation
por: Saijo, Kohei, et al.
Publicado: (2024)
por: Saijo, Kohei, et al.
Publicado: (2024)
Task-Aware Unified Source Separation
por: Saijo, Kohei, et al.
Publicado: (2024)
por: Saijo, Kohei, et al.
Publicado: (2024)
SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers
por: Koo, Junghyun, et al.
Publicado: (2024)
por: Koo, Junghyun, et al.
Publicado: (2024)
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
por: Hussein, Amir, et al.
Publicado: (2025)
por: Hussein, Amir, et al.
Publicado: (2025)
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
por: Saijo, Kohei, et al.
Publicado: (2024)
por: Saijo, Kohei, et al.
Publicado: (2024)
Local Density-Based Anomaly Score Normalization for Domain Generalization
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
por: Wilkinghoff, Kevin, et al.
Publicado: (2025)
Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
por: Saijo, Kohei, et al.
Publicado: (2024)
por: Saijo, Kohei, et al.
Publicado: (2024)
DiffAU: Diffusion-Based Ambisonics Upscaling
por: Milstein, Amit, et al.
Publicado: (2025)
por: Milstein, Amit, et al.
Publicado: (2025)
Factorized RVQ-GAN For Disentangled Speech Tokenization
por: Khurana, Sameer, et al.
Publicado: (2025)
por: Khurana, Sameer, et al.
Publicado: (2025)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
por: Gayer, Yhonatan, et al.
Publicado: (2025)
por: Gayer, Yhonatan, et al.
Publicado: (2025)
Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation
por: Wu, Shih-Lun, et al.
Publicado: (2023)
por: Wu, Shih-Lun, et al.
Publicado: (2023)
BRUDEX Database: Binaural Room Impulse Responses with Uniformly Distributed External Microphones
por: Fejgin, Daniel, et al.
Publicado: (2023)
por: Fejgin, Daniel, et al.
Publicado: (2023)
Speech dereverberation constrained on room impulse response characteristics
por: Bahrman, Louis, et al.
Publicado: (2024)
por: Bahrman, Louis, et al.
Publicado: (2024)
30+ Years of Source Separation Research: Achievements and Future Challenges
por: Araki, Shoko, et al.
Publicado: (2025)
por: Araki, Shoko, et al.
Publicado: (2025)
Acoustivision Pro: An Open-Source Interactive Platform for Room Impulse Response Analysis and Acoustic Characterization
por: Goswami, Mandip
Publicado: (2026)
por: Goswami, Mandip
Publicado: (2026)
Mind the Gap: Detecting Cluster Exits for Robust Local Density-Based Score Normalization in Anomalous Sound Detection
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
Gaunt coefficients for complex and real spherical harmonics with applications to spherical array processing and Ambisonics
por: Politis, Archontis
Publicado: (2024)
por: Politis, Archontis
Publicado: (2024)
TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings
por: Boeddeker, Christoph, et al.
Publicado: (2023)
por: Boeddeker, Christoph, et al.
Publicado: (2023)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
por: Gupta, Shubham, et al.
Publicado: (2024)
por: Gupta, Shubham, et al.
Publicado: (2024)
Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation
por: Ratnarajah, Anton, et al.
Publicado: (2026)
por: Ratnarajah, Anton, et al.
Publicado: (2026)
Acoustic Volume Rendering for Neural Impulse Response Fields
por: Lan, Zitong, et al.
Publicado: (2024)
por: Lan, Zitong, et al.
Publicado: (2024)
Fully Reversing the Shoebox Image Source Method: From Impulse Responses to Room Parameters
por: Sprunck, Tom, et al.
Publicado: (2024)
por: Sprunck, Tom, et al.
Publicado: (2024)
Robot Confirmation Generation and Action Planning Using Long-context Q-Former Integrated with Multimodal LLM
por: Hori, Chiori, et al.
Publicado: (2025)
por: Hori, Chiori, et al.
Publicado: (2025)
Mel-Spectrogram Inversion via Alternating Direction Method of Multipliers
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
por: Masuyama, Yoshiki, et al.
Publicado: (2025)
Exploring the Capability of Mamba in Speech Applications
por: Miyazaki, Koichi, et al.
Publicado: (2024)
por: Miyazaki, Koichi, et al.
Publicado: (2024)
Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
por: Masuyama, Yoshiki, et al.
Publicado: (2024)
GLA-Grad: A Griffin-Lim Extended Waveform Generation Diffusion Model
por: Liu, Haocheng, et al.
Publicado: (2024)
por: Liu, Haocheng, et al.
Publicado: (2024)
SpecDiff-GAN: A Spectrally-Shaped Noise Diffusion GAN for Speech and Music Synthesis
por: Baoueb, Teysir, et al.
Publicado: (2024)
por: Baoueb, Teysir, et al.
Publicado: (2024)
Blind Estimation of Sub-band Acoustic Parameters from Ambisonics Recordings using Spectro-Spatial Covariance Features
por: Meng, Hanyu, et al.
Publicado: (2024)
por: Meng, Hanyu, et al.
Publicado: (2024)
Residual Learning for Neural Ambisonics Encoders
por: Deppisch, Thomas, et al.
Publicado: (2026)
por: Deppisch, Thomas, et al.
Publicado: (2026)
Ejemplares similares
-
Physics-Informed Direction-Aware Neural Acoustic Fields
por: Masuyama, Yoshiki, et al.
Publicado: (2025) -
Direction-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses
por: Ick, Christopher, et al.
Publicado: (2025) -
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
por: Masuyama, Yoshiki, et al.
Publicado: (2024) -
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
por: Masuyama, Yoshiki, et al.
Publicado: (2025) -
FasTUSS: Faster Task-Aware Unified Source Separation
por: Paissan, Francesco, et al.
Publicado: (2025)