Feature-based analysis of oral narratives from Afrikaans and isiXhosa children
Fuente:
arXiv
Salvato in:
| Autori principali: | Sharratt, Emma, Smith, Annelien, Louw, Retief, Klop, Daleen, de Wet, Febe, Kamper, Herman |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Automatically assessing oral narratives of Afrikaans and isiXhosa children
di: Louw, Retief, et al.
Pubblicazione: (2025)
di: Louw, Retief, et al.
Pubblicazione: (2025)
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
di: Jacobs, Christiaan, et al.
Pubblicazione: (2025)
di: Jacobs, Christiaan, et al.
Pubblicazione: (2025)
Towards few-shot isolated word reading assessment
di: Smit, Reuben, et al.
Pubblicazione: (2025)
di: Smit, Reuben, et al.
Pubblicazione: (2025)
Spoken Language Modeling with Duration-Penalized Self-Supervised Units
di: Visser, Nicol, et al.
Pubblicazione: (2025)
di: Visser, Nicol, et al.
Pubblicazione: (2025)
Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features
di: van Rensburg, Kyle Janse, et al.
Pubblicazione: (2026)
di: van Rensburg, Kyle Janse, et al.
Pubblicazione: (2026)
Translating speech with just images
di: Oneata, Dan, et al.
Pubblicazione: (2024)
di: Oneata, Dan, et al.
Pubblicazione: (2024)
Disentanglement in a GAN for Unconditional Speech Synthesis
di: Baas, Matthew, et al.
Pubblicazione: (2023)
di: Baas, Matthew, et al.
Pubblicazione: (2023)
Visually grounded few-shot word learning in low-resource settings
di: Nortje, Leanne, et al.
Pubblicazione: (2023)
di: Nortje, Leanne, et al.
Pubblicazione: (2023)
Spoken-Term Discovery using Discrete Speech Units
di: van Niekerk, Benjamin, et al.
Pubblicazione: (2024)
di: van Niekerk, Benjamin, et al.
Pubblicazione: (2024)
Revisiting speech segmentation and lexicon learning with better features
di: Kamper, Herman, et al.
Pubblicazione: (2024)
di: Kamper, Herman, et al.
Pubblicazione: (2024)
Unsupervised lexicon learning from speech is limited by representations rather than clustering
di: Slabbert, Danel, et al.
Pubblicazione: (2025)
di: Slabbert, Danel, et al.
Pubblicazione: (2025)
ZeroSyl: Simple Zero-Resource Syllable Tokenization for Spoken Language Modeling
di: Visser, Nicol, et al.
Pubblicazione: (2026)
di: Visser, Nicol, et al.
Pubblicazione: (2026)
The mutual exclusivity bias of bilingual visually grounded speech models
di: Oneata, Dan, et al.
Pubblicazione: (2025)
di: Oneata, Dan, et al.
Pubblicazione: (2025)
Visually Grounded Speech Models have a Mutual Exclusivity Bias
di: Nortje, Leanne, et al.
Pubblicazione: (2024)
di: Nortje, Leanne, et al.
Pubblicazione: (2024)
Unsupervised Word Discovery: Boundary Detection with Clustering vs. Dynamic Programming
di: Malan, Simon, et al.
Pubblicazione: (2024)
di: Malan, Simon, et al.
Pubblicazione: (2024)
Should Top-Down Clustering Affect Boundaries in Unsupervised Word Discovery?
di: Malan, Simon, et al.
Pubblicazione: (2025)
di: Malan, Simon, et al.
Pubblicazione: (2025)
LinearVC: Linear transformations of self-supervised features through the lens of voice conversion
di: Kamper, Herman, et al.
Pubblicazione: (2025)
di: Kamper, Herman, et al.
Pubblicazione: (2025)
MARS6: A Small and Robust Hierarchical-Codec Text-to-Speech Model
di: Baas, Matthew, et al.
Pubblicazione: (2025)
di: Baas, Matthew, et al.
Pubblicazione: (2025)
Improved Visually Prompted Keyword Localisation in Real Low-Resource Settings
di: Nortje, Leanne, et al.
Pubblicazione: (2024)
di: Nortje, Leanne, et al.
Pubblicazione: (2024)
Magnitude and Phase-based Feature Fusion Using Co-attention Mechanism for Speaker recognition
di: Su, Rongfeng, et al.
Pubblicazione: (2025)
di: Su, Rongfeng, et al.
Pubblicazione: (2025)
Rhythm Features for Speaker Identification
di: Mehlman, Nick, et al.
Pubblicazione: (2025)
di: Mehlman, Nick, et al.
Pubblicazione: (2025)
BabAR: from phoneme recognition to developmental measures of young children's speech production
di: Lavechin, Marvin, et al.
Pubblicazione: (2026)
di: Lavechin, Marvin, et al.
Pubblicazione: (2026)
Weakly Supervised Phonological Features for Pathological Speech Analysis
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2025)
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2025)
Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis
di: Carbonneau, Marc-André, et al.
Pubblicazione: (2025)
di: Carbonneau, Marc-André, et al.
Pubblicazione: (2025)
Graph-based multi-Feature fusion method for speech emotion recognition
di: Liu, Xueyu, et al.
Pubblicazione: (2024)
di: Liu, Xueyu, et al.
Pubblicazione: (2024)
SCDNet: Self-supervised Learning Feature-based Speaker Change Detection
di: Li, Yue, et al.
Pubblicazione: (2024)
di: Li, Yue, et al.
Pubblicazione: (2024)
On the Role of Spatial Features in Foundation-Model-Based Speaker Diarization
di: Deegen, Marc, et al.
Pubblicazione: (2026)
di: Deegen, Marc, et al.
Pubblicazione: (2026)
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
Exploring Frequency-Domain Feature Modeling for HRTF Magnitude Upsampling
di: Chen, Xingyu, et al.
Pubblicazione: (2026)
di: Chen, Xingyu, et al.
Pubblicazione: (2026)
VAE-based Phoneme Alignment Using Gradient Annealing and SSL Acoustic Features
di: Koriyama, Tomoki
Pubblicazione: (2024)
di: Koriyama, Tomoki
Pubblicazione: (2024)
Audio-Visual Feature Synchronization for Robust Speech Enhancement in Hearing Aids
di: Saleem, Nasir, et al.
Pubblicazione: (2025)
di: Saleem, Nasir, et al.
Pubblicazione: (2025)
Input-Adaptive Spectral Feature Compression by Sequence Modeling for Source Separation
di: Saijo, Kohei, et al.
Pubblicazione: (2026)
di: Saijo, Kohei, et al.
Pubblicazione: (2026)
On HRTF Notch Frequency Prediction Using Anthropometric Features and Neural Networks
di: Arbel, Lior, et al.
Pubblicazione: (2024)
di: Arbel, Lior, et al.
Pubblicazione: (2024)
Distinctive Feature Codec: An Adaptive Efficient Speech Representation for Depression Detection
di: Zhang, Xiangyu, et al.
Pubblicazione: (2025)
di: Zhang, Xiangyu, et al.
Pubblicazione: (2025)
Speech Synthesis From Continuous Features Using Per-Token Latent Diffusion
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
Towards Out-of-Distribution Detection in Vocoder Recognition via Latent Feature Reconstruction
di: Du, Renmingyue, et al.
Pubblicazione: (2024)
di: Du, Renmingyue, et al.
Pubblicazione: (2024)
Latent Secret Spin: Keyed Orthogonal Rotations for Blind Speech Watermarking in Anisotropic Latent Spaces
di: Coletta, Emma, et al.
Pubblicazione: (2026)
di: Coletta, Emma, et al.
Pubblicazione: (2026)
Causal Speech Enhancement with Predicting Semantics based on Quantized Self-supervised Learning Features
di: Tsunoo, Emiru, et al.
Pubblicazione: (2024)
di: Tsunoo, Emiru, et al.
Pubblicazione: (2024)
Why Pre-trained Models Fail: Feature Entanglement in Multi-modal Depression Detection
di: Zhang, Xiangyu, et al.
Pubblicazione: (2025)
di: Zhang, Xiangyu, et al.
Pubblicazione: (2025)
Device Feature based on Graph Fourier Transformation with Logarithmic Processing For Detection of Replay Speech Attacks
di: He, Mingrui, et al.
Pubblicazione: (2024)
di: He, Mingrui, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Automatically assessing oral narratives of Afrikaans and isiXhosa children
di: Louw, Retief, et al.
Pubblicazione: (2025) -
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
di: Jacobs, Christiaan, et al.
Pubblicazione: (2025) -
Towards few-shot isolated word reading assessment
di: Smit, Reuben, et al.
Pubblicazione: (2025) -
Spoken Language Modeling with Duration-Penalized Self-Supervised Units
di: Visser, Nicol, et al.
Pubblicazione: (2025) -
Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features
di: van Rensburg, Kyle Janse, et al.
Pubblicazione: (2026)