EvMic: Event-based Non-contact sound recovery from effective spatial-temporal modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yin, Hao, Guo, Shi, Jia, Xu, XU, Xudong, Zhang, Lu, Liu, Si, Wang, Dong, Lu, Huchuan, Xue, Tianfan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
von: Ma, Wenbo, et al.
Veröffentlicht: (2024)
von: Ma, Wenbo, et al.
Veröffentlicht: (2024)
The UmboMic: A PVDF Cantilever Microphone
von: Yeiser, Aaron J., et al.
Veröffentlicht: (2023)
von: Yeiser, Aaron J., et al.
Veröffentlicht: (2023)
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
von: Liu, Yunyi, et al.
Veröffentlicht: (2025)
von: Liu, Yunyi, et al.
Veröffentlicht: (2025)
Signal processing algorithm effective for sound quality of hearing loss simulators
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
von: Irino, Toshio, et al.
Veröffentlicht: (2024)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
Simi-SFX: A similarity-based conditioning method for controllable sound effect synthesis
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
von: Yue, Haobo, et al.
Veröffentlicht: (2024)
von: Yue, Haobo, et al.
Veröffentlicht: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
von: Fan, Zhiyun, et al.
Veröffentlicht: (2024)
von: Fan, Zhiyun, et al.
Veröffentlicht: (2024)
Differentiable physics for sound field reconstruction
von: Verburg, Samuel A., et al.
Veröffentlicht: (2025)
von: Verburg, Samuel A., et al.
Veröffentlicht: (2025)
Frequency-aware convolution for sound event detection
von: Song, Tao, et al.
Veröffentlicht: (2024)
von: Song, Tao, et al.
Veröffentlicht: (2024)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
von: Xue, Hongfei, et al.
Veröffentlicht: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
The Neural-SRP method for positional sound source localization
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
von: Grinstein, Eric, et al.
Veröffentlicht: (2024)
A Comprehensive Investigation on Speaker Augmentation for Speaker Recognition
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
Lightweight Implicit Neural Network for Binaural Audio Synthesis
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
von: Lu, Xikun, et al.
Veröffentlicht: (2025)
Combining Deterministic Enhanced Conditions with Dual-Streaming Encoding for Diffusion-Based Speech Enhancement
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Some clues to build a sound analysis relevant to hearing
von: Millot, Laurent
Veröffentlicht: (2024)
von: Millot, Laurent
Veröffentlicht: (2024)
Interaural time difference loss for binaural target sound extraction
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
Onset and offset weighted loss function for sound event detection
von: Song, Tao
Veröffentlicht: (2024)
von: Song, Tao
Veröffentlicht: (2024)
Fine-tune the pretrained ATST model for sound event detection
von: Shao, Nian, et al.
Veröffentlicht: (2023)
von: Shao, Nian, et al.
Veröffentlicht: (2023)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
von: Ma, Lu
Veröffentlicht: (2025)
von: Ma, Lu
Veröffentlicht: (2025)
Representational learning for an anomalous sound detection system with source separation model
von: Shin, Seunghyeon, et al.
Veröffentlicht: (2024)
von: Shin, Seunghyeon, et al.
Veröffentlicht: (2024)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
von: Hu, Jinbo, et al.
Veröffentlicht: (2023)
Low-latency Speech Enhancement via Speech Token Generation
von: Xue, Huaying, et al.
Veröffentlicht: (2023)
von: Xue, Huaying, et al.
Veröffentlicht: (2023)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
Convert and Speak: Zero-shot Accent Conversion with Minimum Supervision
von: Jia, Zhijun, et al.
Veröffentlicht: (2024)
von: Jia, Zhijun, et al.
Veröffentlicht: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
The role of direct sound spherical harmonics representation in externalization using binaural reproduction
von: Miller, Eran, et al.
Veröffentlicht: (2024)
von: Miller, Eran, et al.
Veröffentlicht: (2024)
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
von: Faiß, Marius, et al.
Veröffentlicht: (2025)
von: Faiß, Marius, et al.
Veröffentlicht: (2025)
Multispecies bird sound recognition using a fully convolutional neural network
von: García-Ordás, María Teresa, et al.
Veröffentlicht: (2024)
von: García-Ordás, María Teresa, et al.
Veröffentlicht: (2024)
A k-space approach to modeling multi-channel parametric array loudspeaker systems
von: Zhuang, Tao, et al.
Veröffentlicht: (2025)
von: Zhuang, Tao, et al.
Veröffentlicht: (2025)
Efficient learning-based sound propagation for virtual and real-world audio processing applications
von: Ratnarajah, Anton Jeran
Veröffentlicht: (2024)
von: Ratnarajah, Anton Jeran
Veröffentlicht: (2024)
Binaural sound source localization using a hybrid time and frequency domain model
von: Geva, Gil, et al.
Veröffentlicht: (2024)
von: Geva, Gil, et al.
Veröffentlicht: (2024)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
von: Ronchini, Francesca, et al.
Veröffentlicht: (2023)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
von: Ma, Wenbo, et al.
Veröffentlicht: (2024) -
The UmboMic: A PVDF Cantilever Microphone
von: Yeiser, Aaron J., et al.
Veröffentlicht: (2023) -
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
von: Liu, Yunyi, et al.
Veröffentlicht: (2025) -
Signal processing algorithm effective for sound quality of hearing loss simulators
von: Irino, Toshio, et al.
Veröffentlicht: (2024) -
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)