Reduce Computational Complexity for Continuous Wavelet Transform in Acoustic Recognition Using Hop Size
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Phan, Dang Thoai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimal Scalogram for Computational Complexity Reduction in Acoustic Recognition Using Deep Learning
von: Phan, Dang Thoai, et al.
Veröffentlicht: (2025)
von: Phan, Dang Thoai, et al.
Veröffentlicht: (2025)
Comparison Performance of Spectrogram and Scalogram as Input of Acoustic Recognition Task
von: Phan, Dang Thoai
Veröffentlicht: (2024)
von: Phan, Dang Thoai
Veröffentlicht: (2024)
Unsupervised Variational Acoustic Clustering
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025)
EchoScan: Scanning Complex Room Geometries via Acoustic Echoes
von: Yeon, Inmo, et al.
Veröffentlicht: (2023)
von: Yeon, Inmo, et al.
Veröffentlicht: (2023)
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
Cyclic Multichannel Wiener Filter for Acoustic Beamforming
von: Bologni, Giovanni, et al.
Veröffentlicht: (2025)
von: Bologni, Giovanni, et al.
Veröffentlicht: (2025)
Binaural Speech Enhancement Using Complex Convolutional Recurrent Networks
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
WST-X Series: Wavelet Scattering Transform for Interpretable Speech Deepfake Detection
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
von: Xuan, Xi, et al.
Veröffentlicht: (2026)
Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation
von: Kim, Miseul, et al.
Veröffentlicht: (2024)
von: Kim, Miseul, et al.
Veröffentlicht: (2024)
A Speech Production Model for Radar: Connecting Speech Acoustics with Radar-Measured Vibrations
von: Lenz, Isabella, et al.
Veröffentlicht: (2025)
von: Lenz, Isabella, et al.
Veröffentlicht: (2025)
Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models
von: Zucatelli, Guilherme, et al.
Veröffentlicht: (2025)
von: Zucatelli, Guilherme, et al.
Veröffentlicht: (2025)
Graph-Enhanced Dual-Stream Feature Fusion with Pre-Trained Model for Acoustic Traffic Monitoring
von: Fan, Shitong, et al.
Veröffentlicht: (2024)
von: Fan, Shitong, et al.
Veröffentlicht: (2024)
Frequency-Modulated and Single-Tone Excitation to Reveal Vibro-Acoustic Nonlinearities in Loosened Bolted Joints
von: Kullukcu, Berkay, et al.
Veröffentlicht: (2026)
von: Kullukcu, Berkay, et al.
Veröffentlicht: (2026)
Physics-Informed Neural Network-Driven Sparse Field Discretization Method for Near-Field Acoustic Holography
von: Luan, Xinmeng, et al.
Veröffentlicht: (2025)
von: Luan, Xinmeng, et al.
Veröffentlicht: (2025)
Differentiable Acoustic Radiance Transfer
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
Speech-Declipping Transformer with Complex Spectrogram and Learnerble Temporal Features
von: Kwon, Younghoo, et al.
Veröffentlicht: (2024)
von: Kwon, Younghoo, et al.
Veröffentlicht: (2024)
One-Shot Distributed Node-Specific Signal Estimation with Non-Overlapping Latent Subspaces in Acoustic Sensor Networks
von: Didier, Paul, et al.
Veröffentlicht: (2024)
von: Didier, Paul, et al.
Veröffentlicht: (2024)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
von: MacWilliam, Kathleen, et al.
Veröffentlicht: (2024)
von: MacWilliam, Kathleen, et al.
Veröffentlicht: (2024)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
von: Cozens, James M., et al.
Veröffentlicht: (2025)
von: Cozens, James M., et al.
Veröffentlicht: (2025)
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
von: Shore, Noah
Veröffentlicht: (2025)
von: Shore, Noah
Veröffentlicht: (2025)
Heart Murmur and Abnormal PCG Detection via Wavelet Scattering Transform & a 1D-CNN
von: Patwa, Ahmed, et al.
Veröffentlicht: (2023)
von: Patwa, Ahmed, et al.
Veröffentlicht: (2023)
BanglaNum -- A Public Dataset for Bengali Digit Recognition from Speech
von: Mohammad, Mir Sayeed, et al.
Veröffentlicht: (2024)
von: Mohammad, Mir Sayeed, et al.
Veröffentlicht: (2024)
Physics-Informed Direction-Aware Neural Acoustic Fields
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
Machine Learning in Acoustics: A Review and Open-Source Repository
von: McCarthy, Ryan A., et al.
Veröffentlicht: (2025)
von: McCarthy, Ryan A., et al.
Veröffentlicht: (2025)
Classification of Adventitious Sounds Combining Cochleogram and Vision Transformers
von: Mang, Loredana Daria, et al.
Veröffentlicht: (2024)
von: Mang, Loredana Daria, et al.
Veröffentlicht: (2024)
Detection of manatee vocalisations using the Audio Spectrogram Transformer
von: Schiappacasse, Stefano, et al.
Veröffentlicht: (2024)
von: Schiappacasse, Stefano, et al.
Veröffentlicht: (2024)
Scalable-Complexity Steered Response Power Mapping based on Low-Rank and Sparse Interpolation
von: Dietzen, Thomas, et al.
Veröffentlicht: (2023)
von: Dietzen, Thomas, et al.
Veröffentlicht: (2023)
Comparison of Classification Algorithms for COVID19 Detection using Cough Acoustic Signals
von: Erdoğan, Yunus Emre, et al.
Veröffentlicht: (2022)
von: Erdoğan, Yunus Emre, et al.
Veröffentlicht: (2022)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
Semantic Communications for Speech Recognition
von: Weng, Zhenzi, et al.
Veröffentlicht: (2021)
von: Weng, Zhenzi, et al.
Veröffentlicht: (2021)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
Acoustivision Pro: An Open-Source Interactive Platform for Room Impulse Response Analysis and Acoustic Characterization
von: Goswami, Mandip
Veröffentlicht: (2026)
von: Goswami, Mandip
Veröffentlicht: (2026)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
ERes2NetV2: Boosting Short-Duration Speaker Verification Performance with Computational Efficiency
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
von: Chen, Yafeng, et al.
Veröffentlicht: (2024)
On the Invariance of Cross-Correlation Peak Positions Under Monotonic Signal Transformations, with Application to Fast Time Difference Estimation
von: Ueno, Natsuki, et al.
Veröffentlicht: (2025)
von: Ueno, Natsuki, et al.
Veröffentlicht: (2025)
ILD-VIT: A Unified Vision Transformer Architecture for Detection of Interstitial Lung Disease from Respiratory Sounds
von: Hota, Soubhagya Ranjan, et al.
Veröffentlicht: (2025)
von: Hota, Soubhagya Ranjan, et al.
Veröffentlicht: (2025)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
Speakers Localization Using Batch EM In Unfolding Neural Network
von: Veler, Rina, et al.
Veröffentlicht: (2026)
von: Veler, Rina, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Optimal Scalogram for Computational Complexity Reduction in Acoustic Recognition Using Deep Learning
von: Phan, Dang Thoai, et al.
Veröffentlicht: (2025) -
Comparison Performance of Spectrogram and Scalogram as Input of Acoustic Recognition Task
von: Phan, Dang Thoai
Veröffentlicht: (2024) -
Unsupervised Variational Acoustic Clustering
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2025) -
EchoScan: Scanning Complex Room Geometries via Acoustic Echoes
von: Yeon, Inmo, et al.
Veröffentlicht: (2023) -
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)