Quartered Chirp Spectral Envelope for Whispered vs Normal Speech Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Joysingh, S. Johanan, Vijayalakshmi, P., Nagarajan, T. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quartered Spectral Envelope and 1D-CNN-based Classification of Normally Phonated and Whispered Speech
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
Chirp Group Delay based Onset Detection in Instruments with Fast Attack
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
MaskCycleGAN-based Whisper to Normal Speech Conversion
by: Gupta, K. Rohith, et al.
Published: (2024)
by: Gupta, K. Rohith, et al.
Published: (2024)
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
by: Nanmalar, M., et al.
Published: (2024)
by: Nanmalar, M., et al.
Published: (2024)
Development of Large Annotated Music Datasets using HMM-based Forced Viterbi Alignment
by: Joysingh, S. Johanan, et al.
Published: (2024)
by: Joysingh, S. Johanan, et al.
Published: (2024)
Literary and Colloquial Dialect Identification for Tamil using Acoustic Features
by: Nanmalar, M., et al.
Published: (2024)
by: Nanmalar, M., et al.
Published: (2024)
SpecDiff-GAN: A Spectrally-Shaped Noise Diffusion GAN for Speech and Music Synthesis
by: Baoueb, Teysir, et al.
Published: (2024)
by: Baoueb, Teysir, et al.
Published: (2024)
Robustness of Speech Separation Models for Similar-pitch Speakers
by: Lay, Bunlong, et al.
Published: (2024)
by: Lay, Bunlong, et al.
Published: (2024)
Beyond the Labels: Unveiling Text-Dependency in Paralinguistic Speech Recognition Datasets
by: Pešán, Jan, et al.
Published: (2024)
by: Pešán, Jan, et al.
Published: (2024)
DeepFilterGAN: A Full-band Real-time Speech Enhancement System with GAN-based Stochastic Regeneration
by: Serbest, Sanberk, et al.
Published: (2025)
by: Serbest, Sanberk, et al.
Published: (2025)
SpeechMLC: Speech Multi-label Classification
by: Kim, Miseul, et al.
Published: (2025)
by: Kim, Miseul, et al.
Published: (2025)
Speech Watermarking with Discrete Intermediate Representations
by: Ji, Shengpeng, et al.
Published: (2024)
by: Ji, Shengpeng, et al.
Published: (2024)
Self-Tuning Spectral Clustering for Speaker Diarization
by: Raghav, Nikhil, et al.
Published: (2024)
by: Raghav, Nikhil, et al.
Published: (2024)
Gridless Chirp Parameter Retrieval via Constrained Two-Dimensional Atomic Norm Minimization
by: Yang, Dehui, et al.
Published: (2025)
by: Yang, Dehui, et al.
Published: (2025)
TinyChirp: Bird Song Recognition Using TinyML Models on Low-power Wireless Acoustic Sensors
by: Huang, Zhaolan, et al.
Published: (2024)
by: Huang, Zhaolan, et al.
Published: (2024)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
by: Amado-Caballero, Patricia, et al.
Published: (2025)
by: Amado-Caballero, Patricia, et al.
Published: (2025)
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
by: Torres, Bernardo, et al.
Published: (2023)
by: Torres, Bernardo, et al.
Published: (2023)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
by: Baoueb, Teysir, et al.
Published: (2025)
by: Baoueb, Teysir, et al.
Published: (2025)
Speech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
by: Zaiem, Salah, et al.
Published: (2023)
by: Zaiem, Salah, et al.
Published: (2023)
A DNN Based Post-Filter to Enhance the Quality of Coded Speech in MDCT Domain
by: Gupta, Kishan, et al.
Published: (2022)
by: Gupta, Kishan, et al.
Published: (2022)
Lightweight DNN for Full-Band Speech Denoising on Mobile Devices: Exploiting Long and Short Temporal Patterns
by: Drossos, Konstantinos, et al.
Published: (2025)
by: Drossos, Konstantinos, et al.
Published: (2025)
BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing
by: Kawamura, Masaya, et al.
Published: (2025)
by: Kawamura, Masaya, et al.
Published: (2025)
CochCeps-Augment: A Novel Self-Supervised Contrastive Learning Using Cochlear Cepstrum-based Masking for Speech Emotion Recognition
by: Ziogas, Ioannis, et al.
Published: (2024)
by: Ziogas, Ioannis, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
by: Bhattacharjee, Susmita, et al.
Published: (2025)
by: Bhattacharjee, Susmita, et al.
Published: (2025)
Overview of the L3DAS23 Challenge on Audio-Visual Extended Reality
by: Marinoni, Christian, et al.
Published: (2024)
by: Marinoni, Christian, et al.
Published: (2024)
Comparison of Tiny Machine Learning Techniques for Embedded Acoustic Emission Analysis
by: Muthumala, Uditha, et al.
Published: (2024)
by: Muthumala, Uditha, et al.
Published: (2024)
Is Audio Spoof Detection Robust to Laundering Attacks?
by: Ali, Hashim, et al.
Published: (2024)
by: Ali, Hashim, et al.
Published: (2024)
Ultra Low Complexity Deep Learning Based Noise Suppression
by: Shetu, Shrishti Saha, et al.
Published: (2023)
by: Shetu, Shrishti Saha, et al.
Published: (2023)
Anti-aliasing of neural distortion effects via model fine tuning
by: Carson, Alistair, et al.
Published: (2025)
by: Carson, Alistair, et al.
Published: (2025)
Quantifying Articulatory Coordination as a Biomarker for Schizophrenia
by: Premananth, Gowtham, et al.
Published: (2025)
by: Premananth, Gowtham, et al.
Published: (2025)
The ICASSP SP Cadenza Challenge: Music Demixing/Remixing for Hearing Aids
by: Dabike, Gerardo Roa, et al.
Published: (2023)
by: Dabike, Gerardo Roa, et al.
Published: (2023)
Co-Initialization of Control Filter and Secondary Path via Meta-Learning for Active Noise Control
by: Yang, Ziyi, et al.
Published: (2026)
by: Yang, Ziyi, et al.
Published: (2026)
Summary of the DISPLACE Challenge 2023 - DIarization of SPeaker and LAnguage in Conversational Environments
by: Baghel, Shikha, et al.
Published: (2023)
by: Baghel, Shikha, et al.
Published: (2023)
Heart Murmur and Abnormal PCG Detection via Wavelet Scattering Transform & a 1D-CNN
by: Patwa, Ahmed, et al.
Published: (2023)
by: Patwa, Ahmed, et al.
Published: (2023)
Bottleneck Transformer-Based Approach for Improved Automatic STOI Score Prediction
by: Amartyaveer, et al.
Published: (2026)
by: Amartyaveer, et al.
Published: (2026)
The Perception of Phase Intercept Distortion and its Application in Data Augmentation
by: Krishnan, Venkatakrishnan Vaidyanathapuram, et al.
Published: (2025)
by: Krishnan, Venkatakrishnan Vaidyanathapuram, et al.
Published: (2025)
ArrayDPS: Unsupervised Blind Speech Separation with a Diffusion Prior
by: Xu, Zhongweiyang, et al.
Published: (2025)
by: Xu, Zhongweiyang, et al.
Published: (2025)
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
by: Fan, Shitong, et al.
Published: (2024)
by: Fan, Shitong, et al.
Published: (2024)
AudioFuse: Unified Spectral-Temporal Learning via a Hybrid ViT-1D CNN Architecture for Robust Phonocardiogram Classification
by: Siddiqui, Md. Saiful Bari, et al.
Published: (2025)
by: Siddiqui, Md. Saiful Bari, et al.
Published: (2025)
Similar Items
-
Quartered Spectral Envelope and 1D-CNN-based Classification of Normally Phonated and Whispered Speech
by: Joysingh, S. Johanan, et al.
Published: (2024) -
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
by: Joysingh, S. Johanan, et al.
Published: (2024) -
Chirp Group Delay based Onset Detection in Instruments with Fast Attack
by: Joysingh, S. Johanan, et al.
Published: (2024) -
MaskCycleGAN-based Whisper to Normal Speech Conversion
by: Gupta, K. Rohith, et al.
Published: (2024) -
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
by: Nanmalar, M., et al.
Published: (2024)