Using perceptive subbands analysis to perform audio scenes cartography
Fuente:
arXiv
Saved in:
| Main Authors: | Millot, Laurent, Pelé, Gérard, Elliq, Mohammed |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revisiting proximity effect using broadband signals
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
Listening broadband physical model for microphones: a first step
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
Fully Reversing the Shoebox Image Source Method: From Impulse Responses to Room Parameters
by: Sprunck, Tom, et al.
Published: (2024)
by: Sprunck, Tom, et al.
Published: (2024)
Modèle physique variationnel pour l'estimation de réponses impulsionnelles de salles
by: Lalay, Louis, et al.
Published: (2025)
by: Lalay, Louis, et al.
Published: (2025)
Some clues to build a sound analysis relevant to hearing
by: Millot, Laurent
Published: (2024)
by: Millot, Laurent
Published: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
by: Wang, Ziqian, et al.
Published: (2025)
by: Wang, Ziqian, et al.
Published: (2025)
Synthetic training set generation using text-to-audio models for environmental sound classification
by: Ronchini, Francesca, et al.
Published: (2024)
by: Ronchini, Francesca, et al.
Published: (2024)
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
by: Voran, Stephen D.
Published: (2024)
by: Voran, Stephen D.
Published: (2024)
METAMAT 01: A semi-analytic Solution for Benchmarking Wave Propagation Simulations of homogeneous Absorbers in 1D/3D and 2D
by: Schoder, Stefan, et al.
Published: (2024)
by: Schoder, Stefan, et al.
Published: (2024)
Intuitive Control of Scraping and Rubbing Through Audio-tactile Synthesis
by: Aramaki, Mitsuko, et al.
Published: (2024)
by: Aramaki, Mitsuko, et al.
Published: (2024)
Contribution of soundscape appropriateness to soundscape quality assessment in space: a mediating variable affecting acoustic comfort
by: Yang, Xinhao, et al.
Published: (2024)
by: Yang, Xinhao, et al.
Published: (2024)
Correlation and Spectral Density Functions in Mode-Stirred Reverberation -- II. Spectral Moments, Sampling, Noise, EMI and Understirring
by: Arnaut, Luk R., et al.
Published: (2024)
by: Arnaut, Luk R., et al.
Published: (2024)
In situ sound absorption estimation with the discrete complex image source method
by: Brandao, Eric, et al.
Published: (2024)
by: Brandao, Eric, et al.
Published: (2024)
Retrieving Effective Acoustic Impedance and Refractive Index for Size Mismatch Samples
by: Khodaei, Mohammad Javad, et al.
Published: (2021)
by: Khodaei, Mohammad Javad, et al.
Published: (2021)
Compositional nonlinear audio signal processing with Volterra series
by: Araujo-Simon, Jake
Published: (2023)
by: Araujo-Simon, Jake
Published: (2023)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
by: Hiroe, Atsuo, et al.
Published: (2024)
by: Hiroe, Atsuo, et al.
Published: (2024)
Discriminating real and synthetic super-resolved audio samples using embedding-based classifiers
by: Silaev, Mikhail, et al.
Published: (2026)
by: Silaev, Mikhail, et al.
Published: (2026)
Mitigating data replication in text-to-audio generative diffusion models through anti-memorization guidance
by: Messina, Francisco, et al.
Published: (2025)
by: Messina, Francisco, et al.
Published: (2025)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
by: Zhu, Haolin, et al.
Published: (2024)
by: Zhu, Haolin, et al.
Published: (2024)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
by: Shetu, Shrishti Saha, et al.
Published: (2024)
by: Shetu, Shrishti Saha, et al.
Published: (2024)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
by: Shi, Runwu, et al.
Published: (2024)
by: Shi, Runwu, et al.
Published: (2024)
Blind Source Separation of Radar Signals in Time Domain Using Deep Learning
by: Hinderer, Sven
Published: (2025)
by: Hinderer, Sven
Published: (2025)
Optimal Scalogram for Computational Complexity Reduction in Acoustic Recognition Using Deep Learning
by: Phan, Dang Thoai, et al.
Published: (2025)
by: Phan, Dang Thoai, et al.
Published: (2025)
FlowDec: A flow-based full-band general audio codec with high perceptual quality
by: Welker, Simon, et al.
Published: (2025)
by: Welker, Simon, et al.
Published: (2025)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
by: Alimi, Roger, et al.
Published: (2025)
by: Alimi, Roger, et al.
Published: (2025)
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
by: Cheema, Ahsan J., et al.
Published: (2025)
by: Cheema, Ahsan J., et al.
Published: (2025)
Non-locally averaged pruned reassigned spectrograms: a tool for glottal pulse visualization and analysis
by: Griswold, Gabriel J., et al.
Published: (2025)
by: Griswold, Gabriel J., et al.
Published: (2025)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
by: Abu, Avi, et al.
Published: (2024)
by: Abu, Avi, et al.
Published: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
by: Liao, Yuan, et al.
Published: (2025)
by: Liao, Yuan, et al.
Published: (2025)
Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
by: Sechaud, Victor, et al.
Published: (2024)
by: Sechaud, Victor, et al.
Published: (2024)
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
by: Abbott, Leigh, et al.
Published: (2024)
by: Abbott, Leigh, et al.
Published: (2024)
Human-CLAP: Human-perception-based contrastive language-audio pretraining
by: Takano, Taisei, et al.
Published: (2025)
by: Takano, Taisei, et al.
Published: (2025)
Singing Voice Graph Modeling for SingFake Detection
by: Chen, Xuanjun, et al.
Published: (2024)
by: Chen, Xuanjun, et al.
Published: (2024)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
by: Yi, Jayeon, et al.
Published: (2024)
by: Yi, Jayeon, et al.
Published: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
by: Qi, Tianhua, et al.
Published: (2024)
by: Qi, Tianhua, et al.
Published: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
by: Brümann, Klaus, et al.
Published: (2024)
by: Brümann, Klaus, et al.
Published: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
Cross-Talk Reduction
by: Wang, Zhong-Qiu, et al.
Published: (2024)
by: Wang, Zhong-Qiu, et al.
Published: (2024)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
by: Delgado, Pablo M., et al.
Published: (2024)
by: Delgado, Pablo M., et al.
Published: (2024)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
by: Qi, Tianhua, et al.
Published: (2024)
by: Qi, Tianhua, et al.
Published: (2024)
Similar Items
-
Revisiting proximity effect using broadband signals
by: Millot, Laurent, et al.
Published: (2024) -
Listening broadband physical model for microphones: a first step
by: Millot, Laurent, et al.
Published: (2024) -
Fully Reversing the Shoebox Image Source Method: From Impulse Responses to Room Parameters
by: Sprunck, Tom, et al.
Published: (2024) -
Modèle physique variationnel pour l'estimation de réponses impulsionnelles de salles
by: Lalay, Louis, et al.
Published: (2025) -
Some clues to build a sound analysis relevant to hearing
by: Millot, Laurent
Published: (2024)