Gespeichert in:
| Hauptverfasser: | Cozens, James M., Godsill, Simon J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2501.15764 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
von: Shore, Noah
Veröffentlicht: (2025)
von: Shore, Noah
Veröffentlicht: (2025)
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
A Robust Method for Pitch Tracking in the Frequency Following Response using Harmonic Amplitude Summation Filterbank
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
Learnable Adaptive Time-Frequency Representation via Differentiable Short-Time Fourier Transform
von: Leiber, Maxime, et al.
Veröffentlicht: (2025)
von: Leiber, Maxime, et al.
Veröffentlicht: (2025)
BRUDEX Database: Binaural Room Impulse Responses with Uniformly Distributed External Microphones
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Time-domain sound field estimation using kernel ridge regression
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
von: Brunnström, Jesper, et al.
Veröffentlicht: (2025)
LocaGen: Sub-Sample Time-Delay Learning for Beam Localization
von: Kunwar, Ishaan, et al.
Veröffentlicht: (2025)
von: Kunwar, Ishaan, et al.
Veröffentlicht: (2025)
Blind Source Separation of Radar Signals in Time Domain Using Deep Learning
von: Hinderer, Sven
Veröffentlicht: (2025)
von: Hinderer, Sven
Veröffentlicht: (2025)
Note-Level Singing Melody Transcription for Time-Aligned Musical Score Generation
von: Kim, Leekyung, et al.
Veröffentlicht: (2025)
von: Kim, Leekyung, et al.
Veröffentlicht: (2025)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
EMOCONV-DIFF: Diffusion-based Speech Emotion Conversion for Non-parallel and In-the-wild Data
von: Prabhu, Navin Raj, et al.
Veröffentlicht: (2023)
von: Prabhu, Navin Raj, et al.
Veröffentlicht: (2023)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Real-Time Streaming Mel Vocoding with Generative Flow Matching
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
Non-locally averaged pruned reassigned spectrograms: a tool for glottal pulse visualization and analysis
von: Griswold, Gabriel J., et al.
Veröffentlicht: (2025)
von: Griswold, Gabriel J., et al.
Veröffentlicht: (2025)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
MusicHiFi: Fast High-Fidelity Stereo Vocoding
von: Zhu, Ge, et al.
Veröffentlicht: (2024)
von: Zhu, Ge, et al.
Veröffentlicht: (2024)
Learning Perceptually Relevant Temporal Envelope Morphing
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
Benchmarking multi-component signal processing methods in the time-frequency plane
von: Miramont, Juan M., et al.
Veröffentlicht: (2024)
von: Miramont, Juan M., et al.
Veröffentlicht: (2024)
Adaptive Diagonal Loading using Krylov Subspaces for Robust Beamforming
von: Mittal, Manan, et al.
Veröffentlicht: (2026)
von: Mittal, Manan, et al.
Veröffentlicht: (2026)
A Data-Driven Exploration of Elevation Cues in HRTFs: An Explainable AI Perspective Across Multiple Datasets
von: De Rus, Juan Antonio, et al.
Veröffentlicht: (2025)
von: De Rus, Juan Antonio, et al.
Veröffentlicht: (2025)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
von: Monge-Alvarez, Jesús, et al.
Veröffentlicht: (2024)
von: Monge-Alvarez, Jesús, et al.
Veröffentlicht: (2024)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
AutoMashup: Automatic Music Mashups Creation
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
von: Shore, Noah
Veröffentlicht: (2025) -
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024) -
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024) -
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024) -
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2025)