What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training
Fuente:
arXiv
Saved in:
| Main Authors: | Kloots, Marianne de Heer, Mohebbi, Hosein, Pouw, Charlotte, Shen, Gaofei, Zuidema, Willem, Bentum, Martijn |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tracking the emergence of linguistic structure in self-supervised models learning from speech
by: Kloots, Marianne de Heer, et al.
Published: (2026)
by: Kloots, Marianne de Heer, et al.
Published: (2026)
Perception of Phonological Assimilation by Neural Speech Recognition Models
by: Pouw, Charlotte, et al.
Published: (2024)
by: Pouw, Charlotte, et al.
Published: (2024)
In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads
by: Pouw, Charlotte, et al.
Published: (2026)
by: Pouw, Charlotte, et al.
Published: (2026)
Linguists should learn to love speech-based deep learning models
by: Kloots, Marianne de Heer, et al.
Published: (2025)
by: Kloots, Marianne de Heer, et al.
Published: (2025)
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
by: Sauter, Adrian, et al.
Published: (2025)
by: Sauter, Adrian, et al.
Published: (2025)
Exploring bat song syllable representations in self-supervised audio encoders
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
A Linguistically Motivated Analysis of Intonational Phrasing in Text-to-Speech Systems: Revealing Gaps in Syntactic Sensitivity
by: Pouw, Charlotte, et al.
Published: (2025)
by: Pouw, Charlotte, et al.
Published: (2025)
Word stress in self-supervised speech models: A cross-linguistic comparison
by: Bentum, Martijn, et al.
Published: (2025)
by: Bentum, Martijn, et al.
Published: (2025)
On the reliability of feature attribution methods for speech classification
by: Shen, Gaofei, et al.
Published: (2025)
by: Shen, Gaofei, et al.
Published: (2025)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
by: Langedijk, Anna, et al.
Published: (2023)
by: Langedijk, Anna, et al.
Published: (2023)
Disentangling Textual and Acoustic Features of Neural Speech Representations
by: Mohebbi, Hosein, et al.
Published: (2024)
by: Mohebbi, Hosein, et al.
Published: (2024)
Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe
by: Shen, Gaofei, et al.
Published: (2026)
by: Shen, Gaofei, et al.
Published: (2026)
Vision-Language Models Align with Human Neural Representations in Concept Processing
by: Bavaresco, Anna, et al.
Published: (2024)
by: Bavaresco, Anna, et al.
Published: (2024)
Encoding of lexical tone in self-supervised models of spoken language
by: Shen, Gaofei, et al.
Published: (2024)
by: Shen, Gaofei, et al.
Published: (2024)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
Analyzing the relationships between pretraining language, phonetic, tonal, and speaker information in self-supervised speech models
by: Gubian, Michele, et al.
Published: (2025)
by: Gubian, Michele, et al.
Published: (2025)
Self-supervised learning of speech representations with Dutch archival data
by: Vaessen, Nik, et al.
Published: (2025)
by: Vaessen, Nik, et al.
Published: (2025)
AfriHuBERT: A self-supervised speech representation model for African languages
by: Alabi, Jesujoba O., et al.
Published: (2024)
by: Alabi, Jesujoba O., et al.
Published: (2024)
Do self-supervised speech and language models extract similar representations as human brain?
by: Chen, Peili, et al.
Published: (2023)
by: Chen, Peili, et al.
Published: (2023)
How Language Models Prioritize Contextual Grammatical Cues?
by: Amirzadeh, Hamidreza, et al.
Published: (2024)
by: Amirzadeh, Hamidreza, et al.
Published: (2024)
Proboscis apparatus: What do we know (and not know) about nemerteans?
by: Chernyshev, A.V., et al.
Published: (2025)
by: Chernyshev, A.V., et al.
Published: (2025)
On the social bias of speech self-supervised models
by: Lin, Yi-Cheng, et al.
Published: (2024)
by: Lin, Yi-Cheng, et al.
Published: (2024)
Sustainable self-supervised learning for speech representations
by: Lugo, Luis, et al.
Published: (2024)
by: Lugo, Luis, et al.
Published: (2024)
What you should know about seaweeds
Published: (1991)
Published: (1991)
What do you know about siganids?
Published: (1994)
Published: (1994)
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
by: Gowda, Harshavardhana T., et al.
Published: (2025)
by: Gowda, Harshavardhana T., et al.
Published: (2025)
What we know and need to know about Fontan‐associated liver disease
by: Kunimaro Furuta
Published: (2024)
by: Kunimaro Furuta
Published: (2024)
Masked self‐supervised pre‐training model for EEG‐based emotion recognition
by: Xinrong Hu, et al.
Published: (2024)
by: Xinrong Hu, et al.
Published: (2024)
Ensemble of pre-trained language models and data augmentation for hate speech detection from Arabic tweets
by: Daouadi, Kheir Eddine, et al.
Published: (2024)
by: Daouadi, Kheir Eddine, et al.
Published: (2024)
Between surveillance and self‐surveillance: What institutionalised girls in Ciudad Juárez (reveal that they) know about sexuality
by: Bruna Alvarez, et al.
Published: (2024)
by: Bruna Alvarez, et al.
Published: (2024)
Aspects of speech fluency in children with specific language impairment
by: Cláudia Regina Furquim de Andrade
Published: (2014)
by: Cláudia Regina Furquim de Andrade
Published: (2014)
What do we know about Wisconsin lichens?
by: Bennett, James P.
Published: (2006)
by: Bennett, James P.
Published: (2006)
What do we know about the confinement mechanism?
by: Dehghan, Zeinab, et al.
Published: (2024)
by: Dehghan, Zeinab, et al.
Published: (2024)
What do children with cancer know about their medications?
by: Tamara MACDONALD
Published: (2011)
by: Tamara MACDONALD
Published: (2011)
What an anesthesiologist should know about pediatric arrhythmias
by: Michael T. Kuntz, et al.
Published: (2024)
by: Michael T. Kuntz, et al.
Published: (2024)
What do teachers need to know about neuroscience?
by: Michael S. C. Thomas
Published: (2024)
by: Michael S. C. Thomas
Published: (2024)
Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis
by: Huang, De-Xing, et al.
Published: (2026)
by: Huang, De-Xing, et al.
Published: (2026)
Improving conversion rate prediction via self-supervised pre-training in online advertising
by: Shtoff, Alex, et al.
Published: (2024)
by: Shtoff, Alex, et al.
Published: (2024)
Probing self-attention in self-supervised speech models for cross-linguistic differences
by: Gopinath, Sai, et al.
Published: (2024)
by: Gopinath, Sai, et al.
Published: (2024)
Similar Items
-
Tracking the emergence of linguistic structure in self-supervised models learning from speech
by: Kloots, Marianne de Heer, et al.
Published: (2026) -
Perception of Phonological Assimilation by Neural Speech Recognition Models
by: Pouw, Charlotte, et al.
Published: (2024) -
In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads
by: Pouw, Charlotte, et al.
Published: (2026) -
Linguists should learn to love speech-based deep learning models
by: Kloots, Marianne de Heer, et al.
Published: (2025) -
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)