Revisiting MFCCs: Evidence for Spectral-Prosodic Coupling
Fuente:
arXiv
Saved in:
| Main Authors: | Bezerra, Vitor Magno de O. S., Bastos, Gabriel F. A., Montalvão, Jugurta |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Prosodically Informed Mizo TTS without Explicit Tone Markings
by: Mohanta, Abhijit, et al.
Published: (2026)
by: Mohanta, Abhijit, et al.
Published: (2026)
AutoProsody: A Prosodic Feature Extraction Tool for Indian Languages
by: Thinakaran, Preethi, et al.
Published: (2025)
by: Thinakaran, Preethi, et al.
Published: (2025)
Cross-linguistic Prosodic Analysis of Autistic and Non-autistic Child Speech in Finnish, French and Slovak
by: Myllylä, Ida-Lotta, et al.
Published: (2026)
by: Myllylä, Ida-Lotta, et al.
Published: (2026)
ProSDD: Learning Prosodic Representations for Speech Deepfake Detection against Expressive and Emotional Attacks
by: Mahapatra, Aurosweta, et al.
Published: (2026)
by: Mahapatra, Aurosweta, et al.
Published: (2026)
Prosodic Parameter Manipulation in TTS generated speech for Controlled Speech Generation
by: Chary, Podakanti Satyajith
Published: (2024)
by: Chary, Podakanti Satyajith
Published: (2024)
Deep Neural Network for Musical Instrument Recognition using MFCCs
by: Mahanta, Saranga Kingkor, et al.
Published: (2021)
by: Mahanta, Saranga Kingkor, et al.
Published: (2021)
DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions
by: Chen, Weidong, et al.
Published: (2025)
by: Chen, Weidong, et al.
Published: (2025)
PSST! Prosodic Speech Segmentation with Transformers
by: Roll, Nathan, et al.
Published: (2023)
by: Roll, Nathan, et al.
Published: (2023)
Which Prosodic Features Matter Most for Pragmatics?
by: Ward, Nigel G., et al.
Published: (2024)
by: Ward, Nigel G., et al.
Published: (2024)
Sounding Like a Winner? Prosodic Differences in Post-Match Interviews
by: Kakouros, Sofoklis, et al.
Published: (2025)
by: Kakouros, Sofoklis, et al.
Published: (2025)
Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora
by: Onda, Kentaro, et al.
Published: (2025)
by: Onda, Kentaro, et al.
Published: (2025)
Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information
by: Sanders, Nicholas, et al.
Published: (2025)
by: Sanders, Nicholas, et al.
Published: (2025)
Fine-Tuning Whisper for Inclusive Prosodic Stress Analysis
by: Sohn, Samuel S., et al.
Published: (2025)
by: Sohn, Samuel S., et al.
Published: (2025)
Prosodic ABX: A Language-Agnostic Method for Measuring Prosodic Contrast in Speech Representations
by: Sun, Haitong, et al.
Published: (2026)
by: Sun, Haitong, et al.
Published: (2026)
EmphAssess : a Prosodic Benchmark on Assessing Emphasis Transfer in Speech-to-Speech Models
by: de Seyssel, Maureen, et al.
Published: (2023)
by: de Seyssel, Maureen, et al.
Published: (2023)
Pitch Imperfect: Detecting Audio Deepfakes Through Acoustic Prosodic Analysis
by: Warren, Kevin, et al.
Published: (2025)
by: Warren, Kevin, et al.
Published: (2025)
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
by: Ohnaka, Hien, et al.
Published: (2025)
by: Ohnaka, Hien, et al.
Published: (2025)
Modeling Sarcastic Speech: Semantic and Prosodic Cues in a Speech Synthesis Framework
by: Li, Zhu, et al.
Published: (2025)
by: Li, Zhu, et al.
Published: (2025)
A Functional Trade-off between Prosodic and Semantic Cues in Conveying Sarcasm
by: Li, Zhu, et al.
Published: (2024)
by: Li, Zhu, et al.
Published: (2024)
On the application of Visibility Graphs in the Spectral Domain for Speaker Recognition
by: Bocaccio, Hernan, et al.
Published: (2025)
by: Bocaccio, Hernan, et al.
Published: (2025)
Prosodic Structure Beyond Lexical Content: A Study of Self-Supervised Learning
by: Wallbridge, Sarenne, et al.
Published: (2025)
by: Wallbridge, Sarenne, et al.
Published: (2025)
EmergentTTS-Eval: Evaluating TTS Models on Complex Prosodic, Expressiveness, and Linguistic Challenges Using Model-as-a-Judge
by: Manku, Ruskin Raj, et al.
Published: (2025)
by: Manku, Ruskin Raj, et al.
Published: (2025)
Lessons Learnt: Revisit Key Training Strategies for Effective Speech Emotion Recognition in the Wild
by: Tzeng, Jing-Tong, et al.
Published: (2025)
by: Tzeng, Jing-Tong, et al.
Published: (2025)
Revisiting Modeling and Evaluation Approaches in Speech Emotion Recognition: Considering Subjectivity of Annotators and Ambiguity of Emotions
by: Chou, Huang-Cheng, et al.
Published: (2025)
by: Chou, Huang-Cheng, et al.
Published: (2025)
Input-Adaptive Spectral Feature Compression by Sequence Modeling for Source Separation
by: Saijo, Kohei, et al.
Published: (2026)
by: Saijo, Kohei, et al.
Published: (2026)
Revisiting the Privacy of Low-Frequency Speech Signals: Exploring Resampling Methods, Evaluation Scenarios, and Speaker Characteristics
by: Pohlhausen, Jule, et al.
Published: (2025)
by: Pohlhausen, Jule, et al.
Published: (2025)
Revisiting proximity effect using broadband signals
by: Millot, Laurent, et al.
Published: (2024)
by: Millot, Laurent, et al.
Published: (2024)
Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs
by: Langman, Ryan, et al.
Published: (2024)
by: Langman, Ryan, et al.
Published: (2024)
Differentiable Grouped Feedback Delay Networks for Learning Coupled Volume Acoustics
by: Das, Orchisama, et al.
Published: (2025)
by: Das, Orchisama, et al.
Published: (2025)
CFMDCTCodec: A Low-Bitrate Neural Speech Codec with Noise-Prior-aware Conditional Flow Matching for MDCT-Spectral Enhancement
by: Jiang, Xiao-Hang, et al.
Published: (2026)
by: Jiang, Xiao-Hang, et al.
Published: (2026)
Understanding the strengths and weaknesses of SSL models for audio deepfake model attribution
by: Pîrlogeanu, Gabriel, et al.
Published: (2026)
by: Pîrlogeanu, Gabriel, et al.
Published: (2026)
Spectral-Aware Low-Rank Adaptation for Speaker Verification
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Moving Speaker Separation via Parallel Spectral-Spatial Processing
by: Wang, Yuzhu, et al.
Published: (2026)
by: Wang, Yuzhu, et al.
Published: (2026)
The Arrow of Time in Music -- Revisiting the Temporal Structure of Music with Distinguishability and Unique Orientability as the Anchor Point
by: Xu, Qi
Published: (2023)
by: Xu, Qi
Published: (2023)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
by: Eisenberg, Aviad, et al.
Published: (2025)
by: Eisenberg, Aviad, et al.
Published: (2025)
ComFeAT: Combination of Neural and Spectral Features for Improved Depression Detection
by: Phukan, Orchid Chetia, et al.
Published: (2024)
by: Phukan, Orchid Chetia, et al.
Published: (2024)
ESTVocoder: An Excitation-Spectral-Transformed Neural Vocoder Conditioned on Mel Spectrogram
by: Jiang, Xiao-Hang, et al.
Published: (2024)
by: Jiang, Xiao-Hang, et al.
Published: (2024)
Leveraging Joint Spectral and Spatial Learning with MAMBA for Multichannel Speech Enhancement
by: Ren, Wenze, et al.
Published: (2024)
by: Ren, Wenze, et al.
Published: (2024)
Four Decades of Digital Waveguides
by: de Paula, Pablo Tablas, et al.
Published: (2026)
by: de Paula, Pablo Tablas, et al.
Published: (2026)
Similar Items
-
Towards Prosodically Informed Mizo TTS without Explicit Tone Markings
by: Mohanta, Abhijit, et al.
Published: (2026) -
AutoProsody: A Prosodic Feature Extraction Tool for Indian Languages
by: Thinakaran, Preethi, et al.
Published: (2025) -
Cross-linguistic Prosodic Analysis of Autistic and Non-autistic Child Speech in Finnish, French and Slovak
by: Myllylä, Ida-Lotta, et al.
Published: (2026) -
ProSDD: Learning Prosodic Representations for Speech Deepfake Detection against Expressive and Emotional Attacks
by: Mahapatra, Aurosweta, et al.
Published: (2026) -
Prosodic Parameter Manipulation in TTS generated speech for Controlled Speech Generation
by: Chary, Podakanti Satyajith
Published: (2024)