Acoustic characterization of speech rhythm: going beyond metrics with recurrent neural networks
Fuente:
arXiv
Saved in:
| Main Authors: | Deloche, François, Bonnasse-Gahot, Laurent, Gervain, Judit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
by: Sasindran, Zitha, et al.
Published: (2024)
by: Sasindran, Zitha, et al.
Published: (2024)
Denoising by neural network for muzzle blast detection
by: Pujol, Hadrien, et al.
Published: (2025)
by: Pujol, Hadrien, et al.
Published: (2025)
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
by: Kim, Yunsik, et al.
Published: (2025)
by: Kim, Yunsik, et al.
Published: (2025)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
by: Maiti, Soumi, et al.
Published: (2023)
by: Maiti, Soumi, et al.
Published: (2023)
Selfsupervised learning for pathological speech detection
by: Sheikh, Shakeel Ahmad
Published: (2024)
by: Sheikh, Shakeel Ahmad
Published: (2024)
Towards the Synthesis of Non-speech Vocalizations
by: Hoq, Enjamamul, et al.
Published: (2024)
by: Hoq, Enjamamul, et al.
Published: (2024)
U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation
by: Dang, Shaoxiang, et al.
Published: (2024)
by: Dang, Shaoxiang, et al.
Published: (2024)
Validation of artificial neural networks to model the acoustic behaviour of induction motors
by: Jimenez-Romero, F. J., et al.
Published: (2024)
by: Jimenez-Romero, F. J., et al.
Published: (2024)
A multimodal dynamical variational autoencoder for audiovisual speech representation learning
by: Sadok, Samir, et al.
Published: (2023)
by: Sadok, Samir, et al.
Published: (2023)
Unsupervised Acoustic Scene Mapping Based on Acoustic Features and Dimensionality Reduction
by: Cohen, Idan, et al.
Published: (2023)
by: Cohen, Idan, et al.
Published: (2023)
Determining the severity of Parkinson's disease in patients using a multi task neural network
by: García-Ordás, María Teresa, et al.
Published: (2024)
by: García-Ordás, María Teresa, et al.
Published: (2024)
CR-CTC: Consistency regularization on CTC for improved speech recognition
by: Yao, Zengwei, et al.
Published: (2024)
by: Yao, Zengwei, et al.
Published: (2024)
Single-channel speech enhancement using learnable loss mixup
by: Chang, Oscar, et al.
Published: (2023)
by: Chang, Oscar, et al.
Published: (2023)
Zipformer: A faster and better encoder for automatic speech recognition
by: Yao, Zengwei, et al.
Published: (2023)
by: Yao, Zengwei, et al.
Published: (2023)
Robustifying automatic speech recognition by extracting slowly varying features
by: Pizarro, Matías, et al.
Published: (2021)
by: Pizarro, Matías, et al.
Published: (2021)
Late fusion ensembles for speech recognition on diverse input audio representations
by: Jezidžić, Marin, et al.
Published: (2024)
by: Jezidžić, Marin, et al.
Published: (2024)
Boosting keyword spotting through on-device learnable user speech characteristics
by: Cioflan, Cristian, et al.
Published: (2024)
by: Cioflan, Cristian, et al.
Published: (2024)
Generalizable speech deepfake detection via meta-learned LoRA
by: Laakkonen, Janne, et al.
Published: (2025)
by: Laakkonen, Janne, et al.
Published: (2025)
Context-aware child-directed speech detection from long-form recordings
by: Charlot, Théo, et al.
Published: (2026)
by: Charlot, Théo, et al.
Published: (2026)
Dementia classification from spontaneous speech using wrapper-based feature selection
by: Niemelä, Marko, et al.
Published: (2025)
by: Niemelä, Marko, et al.
Published: (2025)
Acoustics-specific Piano Velocity Estimation
by: Simonetta, Federico, et al.
Published: (2022)
by: Simonetta, Federico, et al.
Published: (2022)
Deep Feature Learning for Medical Acoustics
by: Poirè, Alessandro Maria, et al.
Published: (2022)
by: Poirè, Alessandro Maria, et al.
Published: (2022)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
by: Pepino, Leonardo, et al.
Published: (2024)
by: Pepino, Leonardo, et al.
Published: (2024)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
by: Fernández-Díaz, Miguel, et al.
Published: (2024)
by: Fernández-Díaz, Miguel, et al.
Published: (2024)
Benchmarking Representations for Speech, Music, and Acoustic Events
by: La Quatra, Moreno, et al.
Published: (2024)
by: La Quatra, Moreno, et al.
Published: (2024)
Autoregressive Speech Enhancement via Acoustic Tokens
by: Della Libera, Luca, et al.
Published: (2025)
by: Della Libera, Luca, et al.
Published: (2025)
Toward Faithful Explanations in Acoustic Anomaly Detection
by: Elrashid, Maab, et al.
Published: (2026)
by: Elrashid, Maab, et al.
Published: (2026)
CognoSpeak: an automatic, remote assessment of early cognitive decline in real-world conversational speech
by: Pahar, Madhurananda, et al.
Published: (2025)
by: Pahar, Madhurananda, et al.
Published: (2025)
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
by: Pezzoli, Mirco, et al.
Published: (2023)
by: Pezzoli, Mirco, et al.
Published: (2023)
Acoustic identification of individual animals with hierarchical contrastive learning
by: Nolasco, Ines, et al.
Published: (2024)
by: Nolasco, Ines, et al.
Published: (2024)
Underwater Acoustic Signal Recognition Based on Salient Feature
by: Chen, Minghao
Published: (2023)
by: Chen, Minghao
Published: (2023)
Acoustic Classification of Maritime Vessels using Learnable Filterbanks
by: Elsborg, Jonas, et al.
Published: (2025)
by: Elsborg, Jonas, et al.
Published: (2025)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
by: Maisonneuve, Malo, et al.
Published: (2024)
by: Maisonneuve, Malo, et al.
Published: (2024)
Objective and subjective evaluation of speech enhancement methods in the UDASE task of the 7th CHiME challenge
by: Leglaive, Simon, et al.
Published: (2024)
by: Leglaive, Simon, et al.
Published: (2024)
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
by: Agashe, Atharva, et al.
Published: (2025)
by: Agashe, Atharva, et al.
Published: (2025)
Multi-Channel Replay Speech Detection using Acoustic Maps
by: Neri, Michael, et al.
Published: (2026)
by: Neri, Michael, et al.
Published: (2026)
Advancing Test-Time Adaptation in Wild Acoustic Test Settings
by: Liu, Hongfu, et al.
Published: (2023)
by: Liu, Hongfu, et al.
Published: (2023)
DualSep: A Light-weight dual-encoder convolutional recurrent network for real-time in-car speech separation
by: Wang, Ziqian, et al.
Published: (2024)
by: Wang, Ziqian, et al.
Published: (2024)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
by: Ma, Lu
Published: (2025)
by: Ma, Lu
Published: (2025)
Similar Items
-
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
by: Maharana, Sarthak Kumar, et al.
Published: (2023) -
SeMaScore : a new evaluation metric for automatic speech recognition tasks
by: Sasindran, Zitha, et al.
Published: (2024) -
Denoising by neural network for muzzle blast detection
by: Pujol, Hadrien, et al.
Published: (2025) -
Throat and acoustic paired speech dataset for deep learning-based speech enhancement
by: Kim, Yunsik, et al.
Published: (2025) -
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
by: Maiti, Soumi, et al.
Published: (2023)