Gespeichert in:
| Hauptverfasser: | Yu, Haocheng, Ahuja, Krishan K., Sankar, Lakshmi N., Bryngelson, Spencer H. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.16355 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Water Flow Detection Device Based on Sound Data Analysis and Machine Learning to Detect Water Leakage
von: Pourmehrani, Hossein, et al.
Veröffentlicht: (2025)
von: Pourmehrani, Hossein, et al.
Veröffentlicht: (2025)
SonicRAG : High Fidelity Sound Effects Synthesis Based on Retrival Augmented Generation
von: Guo, Yu-Ren, et al.
Veröffentlicht: (2025)
von: Guo, Yu-Ren, et al.
Veröffentlicht: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
Content Leakage in LibriSpeech and Its Impact on the Privacy Evaluation of Speaker Anonymization
von: Franzreb, Carlos, et al.
Veröffentlicht: (2026)
von: Franzreb, Carlos, et al.
Veröffentlicht: (2026)
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)
A Two-Step Learning Framework for Enhancing Sound Event Localization and Detection
von: Yu, Hogeon
Veröffentlicht: (2025)
von: Yu, Hogeon
Veröffentlicht: (2025)
Codec-SUPERB: An In-Depth Analysis of Sound Codec Models
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
von: Wu, Donghang, et al.
Veröffentlicht: (2024)
Sound Zone Control Robust To Sound Speed Change
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Sankha Subhra, et al.
Veröffentlicht: (2024)
AudSemThinker: Enhancing Audio-Language Models through Reasoning over Semantics of Sound
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
von: Wijngaard, Gijs, et al.
Veröffentlicht: (2025)
Evaluating Sound Similarity Metrics for Differentiable, Iterative Sound-Matching
von: Salimi, Amir, et al.
Veröffentlicht: (2025)
von: Salimi, Amir, et al.
Veröffentlicht: (2025)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
von: He, Ke-Xin, et al.
Veröffentlicht: (2019)
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
von: Ai, Yang, et al.
Veröffentlicht: (2024)
von: Ai, Yang, et al.
Veröffentlicht: (2024)
HSDreport: Heart Sound Diagnosis with Echocardiography Reports
von: Zhao, Zihan, et al.
Veröffentlicht: (2024)
von: Zhao, Zihan, et al.
Veröffentlicht: (2024)
Semantic MIMO Systems for Speech-to-Text Transmission
von: Weng, Zhenzi, et al.
Veröffentlicht: (2024)
von: Weng, Zhenzi, et al.
Veröffentlicht: (2024)
Improving Anomalous Sound Detection through Pseudo-anomalous Set Selection and Pseudo-label Utilization under Unlabeled Conditions
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
von: Kuroyanagi, Ibuki, et al.
Veröffentlicht: (2025)
STFTCodec: High-Fidelity Audio Compression through Time-Frequency Domain Representation
von: Feng, Tao, et al.
Veröffentlicht: (2025)
von: Feng, Tao, et al.
Veröffentlicht: (2025)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference Tasks
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
Fractional Fourier Sound Synthesis
von: Gutiérrez, Esteban, et al.
Veröffentlicht: (2025)
von: Gutiérrez, Esteban, et al.
Veröffentlicht: (2025)
Diffuse Sound Field Synthesis
von: Zotter, Franz, et al.
Veröffentlicht: (2024)
von: Zotter, Franz, et al.
Veröffentlicht: (2024)
Sound Event Bounding Boxes
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
von: Ebbers, Janek, et al.
Veröffentlicht: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
SoundBeam meets M2D: Target Sound Extraction with Audio Foundation Model
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
von: Hernandez-Olivan, Carlos, et al.
Veröffentlicht: (2024)
A Generalist Audio Foundation Model for Comprehensive Body Sound Auscultation
von: Wang, Pingjie, et al.
Veröffentlicht: (2024)
von: Wang, Pingjie, et al.
Veröffentlicht: (2024)
SoundLoCD: An Efficient Conditional Discrete Contrastive Latent Diffusion Model for Text-to-Sound Generation
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
A Detailed Audio-Text Data Simulation Pipeline using Single-Event Sounds
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
DiveSound: LLM-Assisted Automatic Taxonomy Construction for Diverse Audio Generation
von: Li, Baihan, et al.
Veröffentlicht: (2024)
von: Li, Baihan, et al.
Veröffentlicht: (2024)
A MATLAB toolbox for Computation of Speech Transmission Index (STI)
von: Rajmic, Pavel, et al.
Veröffentlicht: (2025)
von: Rajmic, Pavel, et al.
Veröffentlicht: (2025)
Fast Algorithm for Moving Sound Source
von: Yang, Dong
Veröffentlicht: (2025)
von: Yang, Dong
Veröffentlicht: (2025)
Boundary-Informed Sound Field Reconstruction
von: Sundström, David, et al.
Veröffentlicht: (2025)
von: Sundström, David, et al.
Veröffentlicht: (2025)
Sound Field Synthesis with Acoustic Waves
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
von: Imoto, Keisuke
Veröffentlicht: (2025)
von: Imoto, Keisuke
Veröffentlicht: (2025)
Two-sided Acoustic Metascreen for Broadband and Individual Reflection and Transmission Control
von: Chen, Ao, et al.
Veröffentlicht: (2024)
von: Chen, Ao, et al.
Veröffentlicht: (2024)
AudioSpa: Spatializing Sound Events with Text
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
Adaptive Differential Denoising for Respiratory Sounds Classification
von: Dong, Gaoyang, et al.
Veröffentlicht: (2025)
von: Dong, Gaoyang, et al.
Veröffentlicht: (2025)
Region-Specific Audio Tagging for Spatial Sound
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
Frequency Dynamic Convolutions for Sound Event Detection
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
Domain-Invariant Representation Learning of Bird Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Water Flow Detection Device Based on Sound Data Analysis and Machine Learning to Detect Water Leakage
von: Pourmehrani, Hossein, et al.
Veröffentlicht: (2025) -
SonicRAG : High Fidelity Sound Effects Synthesis Based on Retrival Augmented Generation
von: Guo, Yu-Ren, et al.
Veröffentlicht: (2025) -
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025) -
Content Leakage in LibriSpeech and Its Impact on the Privacy Evaluation of Speaker Anonymization
von: Franzreb, Carlos, et al.
Veröffentlicht: (2026) -
Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction
von: Lu, Ye-Xin, et al.
Veröffentlicht: (2024)