Disentangled Acoustic Fields For Multimodal Physical Scene Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Jie, Luo, Andrew, Du, Yilun, Cherian, Anoop, Marks, Tim K., Roux, Jonathan Le, Gan, Chuang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
by: Hussein, Amir, et al.
Published: (2025)
by: Hussein, Amir, et al.
Published: (2025)
On Adversarial Attacks In Acoustic Drone Localization
by: Shor, Tamir, et al.
Published: (2025)
by: Shor, Tamir, et al.
Published: (2025)
Soft Acoustic Curvature Sensor: Design and Development
by: Sofla, Mohammad Sheikh, et al.
Published: (2024)
by: Sofla, Mohammad Sheikh, et al.
Published: (2024)
AIM: Acoustic Inertial Measurement for Indoor Drone Localization and Tracking
by: Sun, Yimiao, et al.
Published: (2025)
by: Sun, Yimiao, et al.
Published: (2025)
Physics-Informed Direction-Aware Neural Acoustic Fields
by: Masuyama, Yoshiki, et al.
Published: (2025)
by: Masuyama, Yoshiki, et al.
Published: (2025)
SoundLoc3D: Invisible 3D Sound Source Localization and Classification Using a Multimodal RGB-D Acoustic Camera
by: He, Yuhang, et al.
Published: (2024)
by: He, Yuhang, et al.
Published: (2024)
SonicSense: Object Perception from In-Hand Acoustic Vibration
by: Liu, Jiaxun, et al.
Published: (2024)
by: Liu, Jiaxun, et al.
Published: (2024)
Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
by: Masuyama, Yoshiki, et al.
Published: (2025)
by: Masuyama, Yoshiki, et al.
Published: (2025)
Improving Acoustic Scene Classification in Low-Resource Conditions
by: Chen, Zhi, et al.
Published: (2024)
by: Chen, Zhi, et al.
Published: (2024)
NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
by: Masuyama, Yoshiki, et al.
Published: (2024)
by: Masuyama, Yoshiki, et al.
Published: (2024)
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
by: Yeo, Jin Jie Sean, et al.
Published: (2024)
by: Yeo, Jin Jie Sean, et al.
Published: (2024)
Mind the Gap: Detecting Cluster Exits for Robust Local Density-Based Score Normalization in Anomalous Sound Detection
by: Wilkinghoff, Kevin, et al.
Published: (2026)
by: Wilkinghoff, Kevin, et al.
Published: (2026)
Sound Event Bounding Boxes
by: Ebbers, Janek, et al.
Published: (2024)
by: Ebbers, Janek, et al.
Published: (2024)
Robot Confirmation Generation and Action Planning Using Long-context Q-Former Integrated with Multimodal LLM
by: Hori, Chiori, et al.
Published: (2025)
by: Hori, Chiori, et al.
Published: (2025)
Sound Field Synthesis with Acoustic Waves
by: Mansour, Mohamed F.
Published: (2024)
by: Mansour, Mohamed F.
Published: (2024)
Theoretical Framework for the Optimization of Microphone Array Configuration for Humanoid Robot Audition
by: Tourbabin, Vladimir, et al.
Published: (2024)
by: Tourbabin, Vladimir, et al.
Published: (2024)
Long-Term, Store-Front Robotics: Interactive Music for Robotic Arm, Caxixi and Frame Drums
by: Savery, Richard, et al.
Published: (2024)
by: Savery, Richard, et al.
Published: (2024)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
by: Wang, Jiang, et al.
Published: (2025)
by: Wang, Jiang, et al.
Published: (2025)
Music to Dance as Language Translation using Sequence Models
by: Correia, André, et al.
Published: (2024)
by: Correia, André, et al.
Published: (2024)
SonicBoom: Contact Localization Using Array of Microphones
by: Lee, Moonyoung, et al.
Published: (2024)
by: Lee, Moonyoung, et al.
Published: (2024)
Active Listener: Continuous Generation of Listener's Head Motion Response in Dyadic Interactions
by: Ghosh, Bishal, et al.
Published: (2024)
by: Ghosh, Bishal, et al.
Published: (2024)
Audio Array-Based 3D UAV Trajectory Estimation with LiDAR Pseudo-Labeling
by: Lei, Allen, et al.
Published: (2024)
by: Lei, Allen, et al.
Published: (2024)
Evaluating Speech-in-Speech Perception via a Humanoid Robot
by: Meyer, Luke, et al.
Published: (2023)
by: Meyer, Luke, et al.
Published: (2023)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
by: Lin, Zirui, et al.
Published: (2025)
by: Lin, Zirui, et al.
Published: (2025)
Human-mimetic binaural ear design and sound source direction estimation for task realization of musculoskeletal humanoids
by: Omura, Yusuke, et al.
Published: (2024)
by: Omura, Yusuke, et al.
Published: (2024)
Direction of Arrival Estimation Using Microphone Array Processing for Moving Humanoid Robots
by: Tourbabin, Vladimir, et al.
Published: (2024)
by: Tourbabin, Vladimir, et al.
Published: (2024)
Sim2Real Transfer for Audio-Visual Navigation with Frequency-Adaptive Acoustic Field Prediction
by: Chen, Changan, et al.
Published: (2024)
by: Chen, Changan, et al.
Published: (2024)
Enhanced Reverberation as Supervision for Unsupervised Speech Separation
by: Saijo, Kohei, et al.
Published: (2024)
by: Saijo, Kohei, et al.
Published: (2024)
Task-Aware Unified Source Separation
by: Saijo, Kohei, et al.
Published: (2024)
by: Saijo, Kohei, et al.
Published: (2024)
SMITIN: Self-Monitored Inference-Time INtervention for Generative Music Transformers
by: Koo, Junghyun, et al.
Published: (2024)
by: Koo, Junghyun, et al.
Published: (2024)
TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
by: Saijo, Kohei, et al.
Published: (2024)
by: Saijo, Kohei, et al.
Published: (2024)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
by: Schmid, Florian, et al.
Published: (2025)
by: Schmid, Florian, et al.
Published: (2025)
Low-Complexity Acoustic Scene Classification Using Parallel Attention-Convolution Network
by: Li, Yanxiong, et al.
Published: (2024)
by: Li, Yanxiong, et al.
Published: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
by: Schmid, Florian, et al.
Published: (2024)
by: Schmid, Florian, et al.
Published: (2024)
Online Domain-Incremental Learning Approach to Classify Acoustic Scenes in All Locations
by: Mulimani, Manjunath, et al.
Published: (2024)
by: Mulimani, Manjunath, et al.
Published: (2024)
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
by: Cai, Yiqiang, et al.
Published: (2024)
by: Cai, Yiqiang, et al.
Published: (2024)
Neural Kalman Filters for Acoustic Echo Cancellation
by: Seidel, Ernst, et al.
Published: (2025)
by: Seidel, Ernst, et al.
Published: (2025)
TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings
by: Boeddeker, Christoph, et al.
Published: (2023)
by: Boeddeker, Christoph, et al.
Published: (2023)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
by: Masuyama, Yoshiki, et al.
Published: (2026)
by: Masuyama, Yoshiki, et al.
Published: (2026)
Acoustic Volume Rendering for Neural Impulse Response Fields
by: Lan, Zitong, et al.
Published: (2024)
by: Lan, Zitong, et al.
Published: (2024)
Similar Items
-
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
by: Hussein, Amir, et al.
Published: (2025) -
On Adversarial Attacks In Acoustic Drone Localization
by: Shor, Tamir, et al.
Published: (2025) -
Soft Acoustic Curvature Sensor: Design and Development
by: Sofla, Mohammad Sheikh, et al.
Published: (2024) -
AIM: Acoustic Inertial Measurement for Indoor Drone Localization and Tracking
by: Sun, Yimiao, et al.
Published: (2025) -
Physics-Informed Direction-Aware Neural Acoustic Fields
by: Masuyama, Yoshiki, et al.
Published: (2025)