Point Processes and spatial statistics in time-frequency analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pascal, Barbara, Bardenet, Rémi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Benchmarking multi-component signal processing methods in the time-frequency plane
von: Miramont, Juan M., et al.
Veröffentlicht: (2024)
von: Miramont, Juan M., et al.
Veröffentlicht: (2024)
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
von: Picard, Emilio, et al.
Veröffentlicht: (2026)
von: Picard, Emilio, et al.
Veröffentlicht: (2026)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
von: Haeb-Umbach, Reinhold, et al.
Veröffentlicht: (2025)
von: Haeb-Umbach, Reinhold, et al.
Veröffentlicht: (2025)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
AI-Driven Cardiorespiratory Signal Processing: Separation, Clustering, and Anomaly Detection
von: Torabi, Yasaman
Veröffentlicht: (2026)
von: Torabi, Yasaman
Veröffentlicht: (2026)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
von: Meng, Hanyu, et al.
Veröffentlicht: (2025)
von: Meng, Hanyu, et al.
Veröffentlicht: (2025)
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
von: Abbott, Leigh, et al.
Veröffentlicht: (2024)
von: Abbott, Leigh, et al.
Veröffentlicht: (2024)
DSP-informed bandwidth extension using locally-conditioned excitation and linear time-varying filter subnetworks
von: Nercessian, Shahan, et al.
Veröffentlicht: (2024)
von: Nercessian, Shahan, et al.
Veröffentlicht: (2024)
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
von: Voran, Stephen D.
Veröffentlicht: (2024)
von: Voran, Stephen D.
Veröffentlicht: (2024)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
von: Waheed, Muhammad Fasih, et al.
Veröffentlicht: (2026)
von: Waheed, Muhammad Fasih, et al.
Veröffentlicht: (2026)
Non-locally averaged pruned reassigned spectrograms: a tool for glottal pulse visualization and analysis
von: Griswold, Gabriel J., et al.
Veröffentlicht: (2025)
von: Griswold, Gabriel J., et al.
Veröffentlicht: (2025)
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
Singing Voice Graph Modeling for SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
Cross-Talk Reduction
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Binaural Selective Attention Model for Target Speaker Extraction
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
Constant Directivity Loudspeaker Beamforming
von: Luo, Yuancheng
Veröffentlicht: (2024)
von: Luo, Yuancheng
Veröffentlicht: (2024)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
The CARFAC v2 Cochlear Model in Matlab, NumPy, and JAX
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
What is Learnt by the LEArnable Front-end (LEAF)? Adapting Per-Channel Energy Normalisation (PCEN) to Noisy Conditions
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
Informed FastICA: Semi-Blind Minimum Variance Distortionless Beamformer
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2024)
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2024)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Computational Analysis of Yaredawi YeZema Silt in Ethiopian Orthodox Tewahedo Church Chants
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
Characteristics-Based Design of Generalized-Exponent Bandpass Filters
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
Ähnliche Einträge
-
Benchmarking multi-component signal processing methods in the time-frequency plane
von: Miramont, Juan M., et al.
Veröffentlicht: (2024) -
From the perspective of perceptual speech quality: The robustness of frequency bands to noise
von: Fan, Junyi, et al.
Veröffentlicht: (2025) -
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024) -
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
von: Gupta, Shubham, et al.
Veröffentlicht: (2024) -
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
von: Picard, Emilio, et al.
Veröffentlicht: (2026)