Guardado en:
| Autores principales: | Yin, Hao, Guo, Shi, Jia, Xu, XU, Xudong, Zhang, Lu, Liu, Si, Wang, Dong, Lu, Huchuan, Xue, Tianfan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2504.02402 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The UmboMic: A PVDF Cantilever Microphone
por: Yeiser, Aaron J., et al.
Publicado: (2023)
por: Yeiser, Aaron J., et al.
Publicado: (2023)
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
por: Ma, Wenbo, et al.
Publicado: (2024)
por: Ma, Wenbo, et al.
Publicado: (2024)
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
por: Liu, Yunyi, et al.
Publicado: (2025)
por: Liu, Yunyi, et al.
Publicado: (2025)
Signal processing algorithm effective for sound quality of hearing loss simulators
por: Irino, Toshio, et al.
Publicado: (2024)
por: Irino, Toshio, et al.
Publicado: (2024)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
por: Zhao, PengYuan, et al.
Publicado: (2024)
por: Zhao, PengYuan, et al.
Publicado: (2024)
Simi-SFX: A similarity-based conditioning method for controllable sound effect synthesis
por: Liu, Yunyi, et al.
Publicado: (2024)
por: Liu, Yunyi, et al.
Publicado: (2024)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
por: Gao, Wenmiao, et al.
Publicado: (2025)
por: Gao, Wenmiao, et al.
Publicado: (2025)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
por: Yue, Haobo, et al.
Publicado: (2024)
por: Yue, Haobo, et al.
Publicado: (2024)
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
por: Fan, Zhiyun, et al.
Publicado: (2024)
por: Fan, Zhiyun, et al.
Publicado: (2024)
Differentiable physics for sound field reconstruction
por: Verburg, Samuel A., et al.
Publicado: (2025)
por: Verburg, Samuel A., et al.
Publicado: (2025)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
por: Garcia-Barrios, Guillermo, et al.
Publicado: (2024)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
por: Xue, Hongfei, et al.
Publicado: (2024)
por: Xue, Hongfei, et al.
Publicado: (2024)
Convert and Speak: Zero-shot Accent Conversion with Minimum Supervision
por: Jia, Zhijun, et al.
Publicado: (2024)
por: Jia, Zhijun, et al.
Publicado: (2024)
A Comprehensive Investigation on Speaker Augmentation for Speaker Recognition
por: Zhou, Zhenyu, et al.
Publicado: (2024)
por: Zhou, Zhenyu, et al.
Publicado: (2024)
Frequency-aware convolution for sound event detection
por: Song, Tao, et al.
Publicado: (2024)
por: Song, Tao, et al.
Publicado: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Lightweight Implicit Neural Network for Binaural Audio Synthesis
por: Lu, Xikun, et al.
Publicado: (2025)
por: Lu, Xikun, et al.
Publicado: (2025)
Combining Deterministic Enhanced Conditions with Dual-Streaming Encoding for Diffusion-Based Speech Enhancement
por: Shi, Hao, et al.
Publicado: (2025)
por: Shi, Hao, et al.
Publicado: (2025)
Low-latency Speech Enhancement via Speech Token Generation
por: Xue, Huaying, et al.
Publicado: (2023)
por: Xue, Huaying, et al.
Publicado: (2023)
The Neural-SRP method for positional sound source localization
por: Grinstein, Eric, et al.
Publicado: (2024)
por: Grinstein, Eric, et al.
Publicado: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
Some clues to build a sound analysis relevant to hearing
por: Millot, Laurent
Publicado: (2024)
por: Millot, Laurent
Publicado: (2024)
Interaural time difference loss for binaural target sound extraction
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
Onset and offset weighted loss function for sound event detection
por: Song, Tao
Publicado: (2024)
por: Song, Tao
Publicado: (2024)
Fine-tune the pretrained ATST model for sound event detection
por: Shao, Nian, et al.
Publicado: (2023)
por: Shao, Nian, et al.
Publicado: (2023)
FSD50K-Solo: Automated Curation of Single-Source Sound Events
por: Yang, Ningyuan, et al.
Publicado: (2026)
por: Yang, Ningyuan, et al.
Publicado: (2026)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
por: Hu, Jinbo, et al.
Publicado: (2023)
por: Hu, Jinbo, et al.
Publicado: (2023)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
por: Ma, Lu
Publicado: (2025)
por: Ma, Lu
Publicado: (2025)
Multimodal Consistency-Guided Reference-Free Data Selection for ASR Accent Adaptation
por: Lei, Ligong, et al.
Publicado: (2026)
por: Lei, Ligong, et al.
Publicado: (2026)
A k-space approach to modeling multi-channel parametric array loudspeaker systems
por: Zhuang, Tao, et al.
Publicado: (2025)
por: Zhuang, Tao, et al.
Publicado: (2025)
Exploring Text-Queried Sound Event Detection with Audio Source Separation
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
por: Xiao, Yang, et al.
Publicado: (2024)
por: Xiao, Yang, et al.
Publicado: (2024)
Representational learning for an anomalous sound detection system with source separation model
por: Shin, Seunghyeon, et al.
Publicado: (2024)
por: Shin, Seunghyeon, et al.
Publicado: (2024)
The role of direct sound spherical harmonics representation in externalization using binaural reproduction
por: Miller, Eran, et al.
Publicado: (2024)
por: Miller, Eran, et al.
Publicado: (2024)
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
por: Faiß, Marius, et al.
Publicado: (2025)
por: Faiß, Marius, et al.
Publicado: (2025)
Multispecies bird sound recognition using a fully convolutional neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
FlashAudio: Rectified Flows for Fast and High-Fidelity Text-to-Audio Generation
por: Liu, Huadai, et al.
Publicado: (2024)
por: Liu, Huadai, et al.
Publicado: (2024)
Efficient learning-based sound propagation for virtual and real-world audio processing applications
por: Ratnarajah, Anton Jeran
Publicado: (2024)
por: Ratnarajah, Anton Jeran
Publicado: (2024)
Binaural sound source localization using a hybrid time and frequency domain model
por: Geva, Gil, et al.
Publicado: (2024)
por: Geva, Gil, et al.
Publicado: (2024)
Ejemplares similares
-
The UmboMic: A PVDF Cantilever Microphone
por: Yeiser, Aaron J., et al.
Publicado: (2023) -
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
por: Ma, Wenbo, et al.
Publicado: (2024) -
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
por: Liu, Yunyi, et al.
Publicado: (2025) -
Signal processing algorithm effective for sound quality of hearing loss simulators
por: Irino, Toshio, et al.
Publicado: (2024) -
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
por: Zhao, PengYuan, et al.
Publicado: (2024)