Benchmarking multi-component signal processing methods in the time-frequency plane
Fuente:
arXiv
Salvato in:
| Autori principali: | Miramont, Juan M., Bardenet, Rémi, Chainais, Pierre, Auger, Francois |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Point Processes and spatial statistics in time-frequency analysis
di: Pascal, Barbara, et al.
Pubblicazione: (2024)
di: Pascal, Barbara, et al.
Pubblicazione: (2024)
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
di: Voran, Stephen D.
Pubblicazione: (2024)
di: Voran, Stephen D.
Pubblicazione: (2024)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
di: Waheed, Muhammad Fasih, et al.
Pubblicazione: (2026)
di: Waheed, Muhammad Fasih, et al.
Pubblicazione: (2026)
Audio signal interpolation using optimal transportation of spectrograms
di: Valdivia, David, et al.
Pubblicazione: (2025)
di: Valdivia, David, et al.
Pubblicazione: (2025)
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
di: Fan, Junyi, et al.
Pubblicazione: (2025)
di: Fan, Junyi, et al.
Pubblicazione: (2025)
Deep learning classification system for coconut maturity levels based on acoustic signals
di: Caladcad, June Anne, et al.
Pubblicazione: (2024)
di: Caladcad, June Anne, et al.
Pubblicazione: (2024)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
Compositional nonlinear audio signal processing with Volterra series
di: Araujo-Simon, Jake
Pubblicazione: (2023)
di: Araujo-Simon, Jake
Pubblicazione: (2023)
Physics-Informed Direction-Aware Neural Acoustic Fields
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2026)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2026)
FasTUSS: Faster Task-Aware Unified Source Separation
di: Paissan, Francesco, et al.
Pubblicazione: (2025)
di: Paissan, Francesco, et al.
Pubblicazione: (2025)
FlexIO: Flexible Single- and Multi-Channel Speech Separation and Enhancement
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2025)
DSP-informed bandwidth extension using locally-conditioned excitation and linear time-varying filter subnetworks
di: Nercessian, Shahan, et al.
Pubblicazione: (2024)
di: Nercessian, Shahan, et al.
Pubblicazione: (2024)
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
di: Sato, Hiroshi, et al.
Pubblicazione: (2024)
di: Sato, Hiroshi, et al.
Pubblicazione: (2024)
MusicHiFi: Fast High-Fidelity Stereo Vocoding
di: Zhu, Ge, et al.
Pubblicazione: (2024)
di: Zhu, Ge, et al.
Pubblicazione: (2024)
A Data-Driven Exploration of Elevation Cues in HRTFs: An Explainable AI Perspective Across Multiple Datasets
di: De Rus, Juan Antonio, et al.
Pubblicazione: (2025)
di: De Rus, Juan Antonio, et al.
Pubblicazione: (2025)
Interfacing PDM MEMS microphones with PFM spiking systems: Application for Neuromorphic Auditory Sensors
di: Jimenez-Fernandez, Angel, et al.
Pubblicazione: (2019)
di: Jimenez-Fernandez, Angel, et al.
Pubblicazione: (2019)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
di: Delgado, Pablo M., et al.
Pubblicazione: (2024)
di: Delgado, Pablo M., et al.
Pubblicazione: (2024)
An incremental algorithm based on multichannel non-negative matrix partial co-factorization for ambient denoising in auscultation
di: Cruz, Juan De La Torre, et al.
Pubblicazione: (2024)
di: Cruz, Juan De La Torre, et al.
Pubblicazione: (2024)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
di: Cozens, James M., et al.
Pubblicazione: (2025)
di: Cozens, James M., et al.
Pubblicazione: (2025)
Learning Perceptually Relevant Temporal Envelope Morphing
di: Dixit, Satvik, et al.
Pubblicazione: (2025)
di: Dixit, Satvik, et al.
Pubblicazione: (2025)
Adaptive Diagonal Loading using Krylov Subspaces for Robust Beamforming
di: Mittal, Manan, et al.
Pubblicazione: (2026)
di: Mittal, Manan, et al.
Pubblicazione: (2026)
Gaunt coefficients for complex and real spherical harmonics with applications to spherical array processing and Ambisonics
di: Politis, Archontis
Pubblicazione: (2024)
di: Politis, Archontis
Pubblicazione: (2024)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
di: Monge-Alvarez, Jesús, et al.
Pubblicazione: (2024)
di: Monge-Alvarez, Jesús, et al.
Pubblicazione: (2024)
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
Singing Voice Graph Modeling for SingFake Detection
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
di: Yi, Jayeon, et al.
Pubblicazione: (2024)
di: Yi, Jayeon, et al.
Pubblicazione: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
di: Qi, Tianhua, et al.
Pubblicazione: (2024)
di: Qi, Tianhua, et al.
Pubblicazione: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
di: Brümann, Klaus, et al.
Pubblicazione: (2024)
di: Brümann, Klaus, et al.
Pubblicazione: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
Cross-Talk Reduction
di: Wang, Zhong-Qiu, et al.
Pubblicazione: (2024)
di: Wang, Zhong-Qiu, et al.
Pubblicazione: (2024)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
di: Qi, Tianhua, et al.
Pubblicazione: (2024)
di: Qi, Tianhua, et al.
Pubblicazione: (2024)
Binaural Selective Attention Model for Target Speaker Extraction
di: Meng, Hanyu, et al.
Pubblicazione: (2024)
di: Meng, Hanyu, et al.
Pubblicazione: (2024)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
di: Hiroe, Atsuo, et al.
Pubblicazione: (2024)
di: Hiroe, Atsuo, et al.
Pubblicazione: (2024)
Constant Directivity Loudspeaker Beamforming
di: Luo, Yuancheng
Pubblicazione: (2024)
di: Luo, Yuancheng
Pubblicazione: (2024)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
di: Kechris, Christodoulos, et al.
Pubblicazione: (2024)
di: Kechris, Christodoulos, et al.
Pubblicazione: (2024)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
di: Shetu, Shrishti Saha, et al.
Pubblicazione: (2024)
di: Shetu, Shrishti Saha, et al.
Pubblicazione: (2024)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
di: Yan, Haoyin, et al.
Pubblicazione: (2024)
di: Yan, Haoyin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Point Processes and spatial statistics in time-frequency analysis
di: Pascal, Barbara, et al.
Pubblicazione: (2024) -
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
di: Voran, Stephen D.
Pubblicazione: (2024) -
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
di: Waheed, Muhammad Fasih, et al.
Pubblicazione: (2026) -
Audio signal interpolation using optimal transportation of spectrograms
di: Valdivia, David, et al.
Pubblicazione: (2025) -
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
di: Fan, Junyi, et al.
Pubblicazione: (2025)