BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
Fuente:
arXiv
Saved in:
| Main Authors: | Lavechin, Marvin, Sy, Yaya, Titeux, Hadrien, Blandón, María Andrea Cruz, Räsänen, Okko, Bredin, Hervé, Dupoux, Emmanuel, Cristia, Alejandrina |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An open-source voice type classifier for child-centered daylong recordings
by: Lavechin, Marvin, et al.
Published: (2020)
by: Lavechin, Marvin, et al.
Published: (2020)
Challenges in Automated Processing of Speech from Child Wearables: The Case of Voice Type Classifier
by: Kunze, Tarek, et al.
Published: (2025)
by: Kunze, Tarek, et al.
Published: (2025)
BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings
by: Charlot, Théo, et al.
Published: (2025)
by: Charlot, Théo, et al.
Published: (2025)
Computational modeling of early language learning from acoustic speech and audiovisual input without linguistic priors
by: Räsänen, Okko
Published: (2026)
by: Räsänen, Okko
Published: (2026)
Fifteen Years of Child-Centered Long-Form Recordings: Promises, Resources, and Remaining Challenges to Validity
by: Peurey, Loann, et al.
Published: (2025)
by: Peurey, Loann, et al.
Published: (2025)
A model of early word acquisition based on realistic-scale audiovisual naming events
by: Khorrami, Khazar, et al.
Published: (2024)
by: Khorrami, Khazar, et al.
Published: (2024)
Classification errors distort findings in automated speech processing: examples and solutions from child-development research
by: Gautheron, Lucas, et al.
Published: (2025)
by: Gautheron, Lucas, et al.
Published: (2025)
Encoding of lexical tone in self-supervised models of spoken language
by: Shen, Gaofei, et al.
Published: (2024)
by: Shen, Gaofei, et al.
Published: (2024)
From perception to production: how acoustic invariance facilitates articulatory learning in a self-supervised vocal imitation model
by: Lavechin, Marvin, et al.
Published: (2025)
by: Lavechin, Marvin, et al.
Published: (2025)
Context-aware child-directed speech detection from long-form recordings
by: Charlot, Théo, et al.
Published: (2026)
by: Charlot, Théo, et al.
Published: (2026)
Age-Dependent Analysis and Stochastic Generation of Child-Directed Speech
by: Räsänen, Okko, et al.
Published: (2024)
by: Räsänen, Okko, et al.
Published: (2024)
Can phones, syllables, and words emerge as side-products of cross-situational audiovisual learning? -- A computational investigation
by: Khorrami, Khazar, et al.
Published: (2021)
by: Khorrami, Khazar, et al.
Published: (2021)
Enabling automatic transcription of child-centered audio recordings from real-world environments
by: Kocharov, Daniil, et al.
Published: (2025)
by: Kocharov, Daniil, et al.
Published: (2025)
Evaluation of Audio-Visual Alignments in Visually Grounded Speech Models
by: Khorrami, Khazar, et al.
Published: (2021)
by: Khorrami, Khazar, et al.
Published: (2021)
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System
by: Khorrami, Khazar, et al.
Published: (2023)
by: Khorrami, Khazar, et al.
Published: (2023)
LongTail-Swap: benchmarking language models' abilities on rare words
by: Algayres, Robin, et al.
Published: (2025)
by: Algayres, Robin, et al.
Published: (2025)
PFML: Self-Supervised Learning of Time-Series Data Without Representation Collapse
by: Vaaras, Einari, et al.
Published: (2024)
by: Vaaras, Einari, et al.
Published: (2024)
Evaluating Interactive 2D Visualization as a Sample Selection Strategy for Biomedical Time-Series Data Annotation
by: Vaaras, Einari, et al.
Published: (2026)
by: Vaaras, Einari, et al.
Published: (2026)
Employing self-supervised learning models for cross-linguistic child speech maturity classification
by: Zhang, Theo, et al.
Published: (2025)
by: Zhang, Theo, et al.
Published: (2025)
Automated Analysis of Naturalistic Recordings in Early Childhood: Applications, Challenges, and Opportunities
by: Li, Jialu, et al.
Published: (2025)
by: Li, Jialu, et al.
Published: (2025)
Out-of-distribution generalisation in spoken language understanding
by: Porjazovski, Dejan, et al.
Published: (2024)
by: Porjazovski, Dejan, et al.
Published: (2024)
On the calibration of powerset speaker diarization models
by: Plaquet, Alexis, et al.
Published: (2024)
by: Plaquet, Alexis, et al.
Published: (2024)
Exploring cognition processes in second language acquisition: the case of cognates and false-friends in EST
by: Pilar Durán Escribano
Published: (2004)
by: Pilar Durán Escribano
Published: (2004)
Third language acquisition
Published: (2021)
Published: (2021)
Integrating Continuous and Binary Relevances in Audio-Text Relevance Learning
by: Xie, Huang, et al.
Published: (2024)
by: Xie, Huang, et al.
Published: (2024)
Text-based Audio Retrieval by Learning from Similarities between Audio Captions
by: Xie, Huang, et al.
Published: (2024)
by: Xie, Huang, et al.
Published: (2024)
Investigating Affect Mining Techniques for Annotation Sample Selection in the Creation of Finnish Affective Speech Corpus
by: Lahtinen, Kalle, et al.
Published: (2025)
by: Lahtinen, Kalle, et al.
Published: (2025)
BabAR: from phoneme recognition to developmental measures of young children's speech production
by: Lavechin, Marvin, et al.
Published: (2026)
by: Lavechin, Marvin, et al.
Published: (2026)
Multilingualism and third language acquisition
Published: (2021)
Published: (2021)
Adversarial synthesis based data-augmentation for code-switched spoken language identification
by: Shastri, Parth, et al.
Published: (2022)
by: Shastri, Parth, et al.
Published: (2022)
Meeting in the middle: Luria´s approach and cognitive approach to spoken language impairment in aphasia
by: E.I. Markashova
Published: (2022)
by: E.I. Markashova
Published: (2022)
Opening the black box of language acquisition
by: Michaud, Jérôme, et al.
Published: (2024)
by: Michaud, Jérôme, et al.
Published: (2024)
Complexity in second language phonology acquisition
by: Ronaldo Mangueira Lima Júnior
Published: (2013)
by: Ronaldo Mangueira Lima Júnior
Published: (2013)
Intuitive physics understanding emerges from self-supervised pretraining on natural videos
by: Garrido, Quentin, et al.
Published: (2025)
by: Garrido, Quentin, et al.
Published: (2025)
Reflections on the connection between computer-assisted language learning and second language acquisition
by: Olmedo Bula Villalobos
Published: (2012)
by: Olmedo Bula Villalobos
Published: (2012)
Gujarati-English Code-Switching Speech Recognition using ensemble prediction of spoken language
by: Sharma, Yash, et al.
Published: (2024)
by: Sharma, Yash, et al.
Published: (2024)
Is teacher’s English good enough?: A case study of Saudi teacher spoken language
by: Maather AlRawi
Published: (2022)
by: Maather AlRawi
Published: (2022)
Towards a motivating language acquisition curriculum
by: Liam Printer
Published: (2024)
by: Liam Printer
Published: (2024)
Complexity and identity reconstruction in second language acquisition
by: Liliane Assis Sade
Published: (2009)
by: Liliane Assis Sade
Published: (2009)
How predictable is language model benchmark performance?
by: Owen, David
Published: (2024)
by: Owen, David
Published: (2024)
Similar Items
-
An open-source voice type classifier for child-centered daylong recordings
by: Lavechin, Marvin, et al.
Published: (2020) -
Challenges in Automated Processing of Speech from Child Wearables: The Case of Voice Type Classifier
by: Kunze, Tarek, et al.
Published: (2025) -
BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings
by: Charlot, Théo, et al.
Published: (2025) -
Computational modeling of early language learning from acoustic speech and audiovisual input without linguistic priors
by: Räsänen, Okko
Published: (2026) -
Fifteen Years of Child-Centered Long-Form Recordings: Promises, Resources, and Remaining Challenges to Validity
by: Peurey, Loann, et al.
Published: (2025)