Saved in:
| Main Authors: | Schiappacasse, Stefano, de Wolff, Taco, Henaut, Yann, Cervera, Regina, Charles, Aviva, Tobar, Felipe |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.18083 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Spectrogram Separation and TDOA Estimation using Optimal Transport
by: Fabiani, Linda, et al.
Published: (2025)
by: Fabiani, Linda, et al.
Published: (2025)
Speech-Declipping Transformer with Complex Spectrogram and Learnerble Temporal Features
by: Kwon, Younghoo, et al.
Published: (2024)
by: Kwon, Younghoo, et al.
Published: (2024)
Learning Temporal Resolution in Spectrogram for Audio Classification
by: Liu, Haohe, et al.
Published: (2022)
by: Liu, Haohe, et al.
Published: (2022)
A Neural Denoising Vocoder for Clean Waveform Generation from Noisy Mel-Spectrogram based on Amplitude and Phase Predictions
by: Du, Hui-Peng, et al.
Published: (2024)
by: Du, Hui-Peng, et al.
Published: (2024)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
by: Alimi, Roger, et al.
Published: (2025)
by: Alimi, Roger, et al.
Published: (2025)
SSM2Mel: State Space Model to Reconstruct Mel Spectrogram from the EEG
by: Fan, Cunhang, et al.
Published: (2025)
by: Fan, Cunhang, et al.
Published: (2025)
Compression Robust Synthetic Speech Detection Using Patched Spectrogram Transformer
by: Yadav, Amit Kumar Singh, et al.
Published: (2024)
by: Yadav, Amit Kumar Singh, et al.
Published: (2024)
Benchmarking Audio Deepfake Detection Robustness in Real-world Communication Scenarios
by: Shi, Haohan, et al.
Published: (2025)
by: Shi, Haohan, et al.
Published: (2025)
Exploring Audio-Visual Information Fusion for Sound Event Localization and Detection In Low-Resource Realistic Scenarios
by: Jiang, Ya, et al.
Published: (2024)
by: Jiang, Ya, et al.
Published: (2024)
Beyond Identity: A Generalizable Approach for Deepfake Audio Detection
by: Ahmadiadli, Yasaman, et al.
Published: (2025)
by: Ahmadiadli, Yasaman, et al.
Published: (2025)
Audio signal interpolation using optimal transportation of spectrograms
by: Valdivia, David, et al.
Published: (2025)
by: Valdivia, David, et al.
Published: (2025)
Spectrogram features for audio and speech analysis
by: McLoughlin, Ian, et al.
Published: (2026)
by: McLoughlin, Ian, et al.
Published: (2026)
SUNAC: Source-aware Unified Neural Audio Codec
by: Aihara, Ryo, et al.
Published: (2025)
by: Aihara, Ryo, et al.
Published: (2025)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
by: Monge-Alvarez, Jesús, et al.
Published: (2024)
by: Monge-Alvarez, Jesús, et al.
Published: (2024)
Perceptually-motivated Spatial Audio Codec for Higher-Order Ambisonics Compression
by: Hold, Christoph, et al.
Published: (2024)
by: Hold, Christoph, et al.
Published: (2024)
Is Audio Spoof Detection Robust to Laundering Attacks?
by: Ali, Hashim, et al.
Published: (2024)
by: Ali, Hashim, et al.
Published: (2024)
Robust Fixed-Filter Sound Zone Control with Audio-Based Position Tracking
by: Bhattacharjee, Sankha Subhra, et al.
Published: (2024)
by: Bhattacharjee, Sankha Subhra, et al.
Published: (2024)
Generative AI in Signal Processing Education: An Audio Foundation Model Based Approach
by: Khan, Muhammad Salman, et al.
Published: (2026)
by: Khan, Muhammad Salman, et al.
Published: (2026)
ILD-VIT: A Unified Vision Transformer Architecture for Detection of Interstitial Lung Disease from Respiratory Sounds
by: Hota, Soubhagya Ranjan, et al.
Published: (2025)
by: Hota, Soubhagya Ranjan, et al.
Published: (2025)
Zero-Bit Transmission of Adaptive Pre- and De-emphasis Filters for Speech and Audio Coding
by: Piralideh, Niloofar Omidi, et al.
Published: (2024)
by: Piralideh, Niloofar Omidi, et al.
Published: (2024)
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
by: Neri, Michael, et al.
Published: (2025)
by: Neri, Michael, et al.
Published: (2025)
FAST: Fast Audio Spectrogram Transformer
by: Naman, Anugunj, et al.
Published: (2025)
by: Naman, Anugunj, et al.
Published: (2025)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
by: Liao, Yuan, et al.
Published: (2025)
by: Liao, Yuan, et al.
Published: (2025)
Parameter-Efficient Transfer Learning of Audio Spectrogram Transformers
by: Cappellazzo, Umberto, et al.
Published: (2023)
by: Cappellazzo, Umberto, et al.
Published: (2023)
Aliasing-Free Neural Audio Synthesis
by: Gu, Yicheng, et al.
Published: (2025)
by: Gu, Yicheng, et al.
Published: (2025)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
by: Miotello, Federico, et al.
Published: (2024)
by: Miotello, Federico, et al.
Published: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
Low-Complexity Neural Wind Noise Reduction for Audio Recordings
by: Eftekhari, Hesam, et al.
Published: (2025)
by: Eftekhari, Hesam, et al.
Published: (2025)
Classification of Adventitious Sounds Combining Cochleogram and Vision Transformers
by: Mang, Loredana Daria, et al.
Published: (2024)
by: Mang, Loredana Daria, et al.
Published: (2024)
A XAI-based Framework for Frequency Subband Characterization of Cough Spectrograms in Chronic Respiratory Disease
by: Amado-Caballero, Patricia, et al.
Published: (2025)
by: Amado-Caballero, Patricia, et al.
Published: (2025)
Interpolation Filter Design for Sample Rate Independent Audio Effect RNNs
by: Carson, Alistair, et al.
Published: (2024)
by: Carson, Alistair, et al.
Published: (2024)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
by: Carson, Alistair, et al.
Published: (2024)
by: Carson, Alistair, et al.
Published: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
by: Wang, Ziqian, et al.
Published: (2025)
by: Wang, Ziqian, et al.
Published: (2025)
Do Music Source Separation Models Preserve Spatial Information in Binaural Audio?
by: Namballa, Richa, et al.
Published: (2025)
by: Namballa, Richa, et al.
Published: (2025)
Reduce Computational Complexity for Continuous Wavelet Transform in Acoustic Recognition Using Hop Size
by: Phan, Dang Thoai
Published: (2024)
by: Phan, Dang Thoai
Published: (2024)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
by: Meng, Hanyu, et al.
Published: (2025)
by: Meng, Hanyu, et al.
Published: (2025)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
by: Yao, Shengshi, et al.
Published: (2025)
by: Yao, Shengshi, et al.
Published: (2025)
Cough-E: A multimodal, privacy-preserving cough detection algorithm for the edge
by: Albini, Stefano, et al.
Published: (2024)
by: Albini, Stefano, et al.
Published: (2024)
A Zero-Shot Physics-Informed Dictionary Learning Approach for Sound Field Reconstruction
by: Damiano, Stefano, et al.
Published: (2024)
by: Damiano, Stefano, et al.
Published: (2024)
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models
by: Yang, Ningyuan, et al.
Published: (2026)
by: Yang, Ningyuan, et al.
Published: (2026)
Similar Items
-
Joint Spectrogram Separation and TDOA Estimation using Optimal Transport
by: Fabiani, Linda, et al.
Published: (2025) -
Speech-Declipping Transformer with Complex Spectrogram and Learnerble Temporal Features
by: Kwon, Younghoo, et al.
Published: (2024) -
Learning Temporal Resolution in Spectrogram for Audio Classification
by: Liu, Haohe, et al.
Published: (2022) -
A Neural Denoising Vocoder for Clean Waveform Generation from Noisy Mel-Spectrogram based on Amplitude and Phase Predictions
by: Du, Hui-Peng, et al.
Published: (2024) -
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
by: Alimi, Roger, et al.
Published: (2025)