Linguists should learn to love speech-based deep learning models
Fuente:
arXiv
Saved in:
| Main Authors: | Kloots, Marianne de Heer, Boersma, Paul, Zuidema, Willem |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training
by: Kloots, Marianne de Heer, et al.
Published: (2025)
by: Kloots, Marianne de Heer, et al.
Published: (2025)
Brain-tuned Speech Models Better Reflect Speech Processing Stages in the Brain
by: Moussa, Omer, et al.
Published: (2025)
by: Moussa, Omer, et al.
Published: (2025)
Towards an End-to-End Framework for Invasive Brain Signal Decoding with Large Language Models
by: Feng, Sheng, et al.
Published: (2024)
by: Feng, Sheng, et al.
Published: (2024)
Using i-vectors for subject-independent cross-session EEG transfer learning
by: Lasko, Jonathan, et al.
Published: (2024)
by: Lasko, Jonathan, et al.
Published: (2024)
Sigma-lognormal modeling of speech
by: Carmona-Duarte, C., et al.
Published: (2024)
by: Carmona-Duarte, C., et al.
Published: (2024)
Prosody of speech production in latent post-stroke aphasia
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
From Spikes to Speech: NeuroVoc -- A Biologically Plausible Vocoder Framework for Auditory Perception and Cochlear Implant Simulation
by: de Nobel, Jacob, et al.
Published: (2025)
by: de Nobel, Jacob, et al.
Published: (2025)
Interpretable Embeddings of Speech Enhance and Explain Brain Encoding Performance of Audio Models
by: Shimizu, Riki, et al.
Published: (2025)
by: Shimizu, Riki, et al.
Published: (2025)
Désentrelacement Fréquentiel Doux pour les Codecs Audio Neuronaux
by: Giniès, Benoît, et al.
Published: (2025)
by: Giniès, Benoît, et al.
Published: (2025)
Soft-Weighted CrossEntropy Loss for Continous Alzheimer's Disease Detection
by: Zhang, Xiaohui, et al.
Published: (2024)
by: Zhang, Xiaohui, et al.
Published: (2024)
Neural2Speech: A Transfer Learning Framework for Neural-Driven Speech Reconstruction
by: Li, Jiawei, et al.
Published: (2023)
by: Li, Jiawei, et al.
Published: (2023)
Tracking the emergence of linguistic structure in self-supervised models learning from speech
by: Kloots, Marianne de Heer, et al.
Published: (2026)
by: Kloots, Marianne de Heer, et al.
Published: (2026)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Mode-conditioned music learning and composition: a spiking neural network inspired by neuroscience and psychology
by: Liang, Qian, et al.
Published: (2024)
by: Liang, Qian, et al.
Published: (2024)
Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy
by: Kopar, Serli, et al.
Published: (2026)
by: Kopar, Serli, et al.
Published: (2026)
Improving Speech Decoding from ECoG with Self-Supervised Pretraining
by: Yuan, Brian A., et al.
Published: (2024)
by: Yuan, Brian A., et al.
Published: (2024)
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
by: Sato, Motoshige, et al.
Published: (2026)
by: Sato, Motoshige, et al.
Published: (2026)
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
by: Geirnaert, Simon, et al.
Published: (2025)
by: Geirnaert, Simon, et al.
Published: (2025)
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models
by: Ciferri, Matteo, et al.
Published: (2024)
by: Ciferri, Matteo, et al.
Published: (2024)
Do self-supervised speech and language models extract similar representations as human brain?
by: Chen, Peili, et al.
Published: (2023)
by: Chen, Peili, et al.
Published: (2023)
Scaling Law in Neural Data: Non-Invasive Speech Decoding with 175 Hours of EEG Data
by: Sato, Motoshige, et al.
Published: (2024)
by: Sato, Motoshige, et al.
Published: (2024)
AADNet: Exploring EEG Spatiotemporal Information for Fast and Accurate Orientation and Timbre Detection of Auditory Attention Based on A Cue-Masked Paradigm
by: Shi, Keren, et al.
Published: (2025)
by: Shi, Keren, et al.
Published: (2025)
Trade-offs between structural richness and communication efficiency in music network representations
by: Rosselló, Lluc Bono, et al.
Published: (2025)
by: Rosselló, Lluc Bono, et al.
Published: (2025)
Towards Dynamic Neural Communication and Speech Neuroprosthesis Based on Viseme Decoding
by: Park, Ji-Ha, et al.
Published: (2025)
by: Park, Ji-Ha, et al.
Published: (2025)
Understanding Auditory Evoked Brain Signal via Physics-informed Embedding Network with Multi-Task Transformer
by: Ma, Wanli, et al.
Published: (2024)
by: Ma, Wanli, et al.
Published: (2024)
R&B -- Rhythm and Brain: Cross-subject Decoding of Music from Human Brain Activity
by: Ferrante, Matteo, et al.
Published: (2024)
by: Ferrante, Matteo, et al.
Published: (2024)
Human Brain Exhibits Distinct Patterns When Listening to Fake Versus Real Audio: Preliminary Evidence
by: Salehi, Mahsa, et al.
Published: (2024)
by: Salehi, Mahsa, et al.
Published: (2024)
Brain2Music: Reconstructing Music from Human Brain Activity
by: Denk, Timo I., et al.
Published: (2023)
by: Denk, Timo I., et al.
Published: (2023)
On the Within-class Variation Issue in Alzheimer's Disease Detection
by: Kang, Jiawen, et al.
Published: (2024)
by: Kang, Jiawen, et al.
Published: (2024)
Towards Homogeneous Lexical Tone Decoding from Heterogeneous Intracranial Recordings
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
A Penny for Your Thoughts: Decoding Speech from Inexpensive Brain Signals
by: Auster, Quentin, et al.
Published: (2025)
by: Auster, Quentin, et al.
Published: (2025)
Two-component spatiotemporal template for activation-inhibition of speech in ECoG
by: Easthope, Eric
Published: (2024)
by: Easthope, Eric
Published: (2024)
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings
by: Hmamouche, Youssef, et al.
Published: (2024)
by: Hmamouche, Youssef, et al.
Published: (2024)
Decoding Selective Auditory Attention to Musical Elements in Ecologically Valid Music Listening
by: Akama, Taketo, et al.
Published: (2025)
by: Akama, Taketo, et al.
Published: (2025)
Condition-Invariant fMRI Decoding of Speech Intelligibility with Deep State Space Model
by: Sung, Ching-Chih, et al.
Published: (2025)
by: Sung, Ching-Chih, et al.
Published: (2025)
Predicting Artificial Neural Network Representations to Learn Recognition Model for Music Identification from Brain Recordings
by: Akama, Taketo, et al.
Published: (2024)
by: Akama, Taketo, et al.
Published: (2024)
Speech language models lack important brain-relevant semantics
by: Oota, Subba Reddy, et al.
Published: (2023)
by: Oota, Subba Reddy, et al.
Published: (2023)
Towards Decoding Brain Activity During Passive Listening of Speech
by: Fodor, Milán András, et al.
Published: (2024)
by: Fodor, Milán András, et al.
Published: (2024)
A computational loudness model for electrical stimulation with cochlear implants
by: Alvarez, Franklin, et al.
Published: (2025)
by: Alvarez, Franklin, et al.
Published: (2025)
Similar Items
-
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024) -
What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training
by: Kloots, Marianne de Heer, et al.
Published: (2025) -
Brain-tuned Speech Models Better Reflect Speech Processing Stages in the Brain
by: Moussa, Omer, et al.
Published: (2025) -
Towards an End-to-End Framework for Invasive Brain Signal Decoding with Large Language Models
by: Feng, Sheng, et al.
Published: (2024) -
Using i-vectors for subject-independent cross-session EEG transfer learning
by: Lasko, Jonathan, et al.
Published: (2024)