Do self-supervised speech and language models extract similar representations as human brain?
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Peili, He, Linyang, Fu, Li, Fan, Lu, Chang, Edward F., Li, Yuanning |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural2Speech: A Transfer Learning Framework for Neural-Driven Speech Reconstruction
by: Li, Jiawei, et al.
Published: (2023)
by: Li, Jiawei, et al.
Published: (2023)
Speech language models lack important brain-relevant semantics
by: Oota, Subba Reddy, et al.
Published: (2023)
by: Oota, Subba Reddy, et al.
Published: (2023)
A computational loudness model for electrical stimulation with cochlear implants
by: Alvarez, Franklin, et al.
Published: (2025)
by: Alvarez, Franklin, et al.
Published: (2025)
Prosody of speech production in latent post-stroke aphasia
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Linguists should learn to love speech-based deep learning models
by: Kloots, Marianne de Heer, et al.
Published: (2025)
by: Kloots, Marianne de Heer, et al.
Published: (2025)
Soft-Weighted CrossEntropy Loss for Continous Alzheimer's Disease Detection
by: Zhang, Xiaohui, et al.
Published: (2024)
by: Zhang, Xiaohui, et al.
Published: (2024)
Cortical Temporal Mismatch Compensation in Bimodal Cochlear Implant Users: Selective Attention Decoding and Pupillometry Study
by: Dolhopiatenko, Hanna, et al.
Published: (2025)
by: Dolhopiatenko, Hanna, et al.
Published: (2025)
Sigma-lognormal modeling of speech
by: Carmona-Duarte, C., et al.
Published: (2024)
by: Carmona-Duarte, C., et al.
Published: (2024)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
From Spikes to Speech: NeuroVoc -- A Biologically Plausible Vocoder Framework for Auditory Perception and Cochlear Implant Simulation
by: de Nobel, Jacob, et al.
Published: (2025)
by: de Nobel, Jacob, et al.
Published: (2025)
Interpretable Embeddings of Speech Enhance and Explain Brain Encoding Performance of Audio Models
by: Shimizu, Riki, et al.
Published: (2025)
by: Shimizu, Riki, et al.
Published: (2025)
Désentrelacement Fréquentiel Doux pour les Codecs Audio Neuronaux
by: Giniès, Benoît, et al.
Published: (2025)
by: Giniès, Benoît, et al.
Published: (2025)
Trade-offs between structural richness and communication efficiency in music network representations
by: Rosselló, Lluc Bono, et al.
Published: (2025)
by: Rosselló, Lluc Bono, et al.
Published: (2025)
AADNet: Exploring EEG Spatiotemporal Information for Fast and Accurate Orientation and Timbre Detection of Auditory Attention Based on A Cue-Masked Paradigm
by: Shi, Keren, et al.
Published: (2025)
by: Shi, Keren, et al.
Published: (2025)
Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models using Minimal Pairs
by: He, Linyang, et al.
Published: (2024)
by: He, Linyang, et al.
Published: (2024)
neuro2voc: Decoding Vocalizations from Neural Activity
by: Gao, Fei
Published: (2025)
by: Gao, Fei
Published: (2025)
Perception of dynamic multi-speaker auditory scenes under different modes of attention
by: Graceffo, Stephanie, et al.
Published: (2025)
by: Graceffo, Stephanie, et al.
Published: (2025)
Utilizing Information Theoretic Approach to Study Cochlear Neural Degeneration
by: Cheema, Ahsan J., et al.
Published: (2025)
by: Cheema, Ahsan J., et al.
Published: (2025)
Two-component spatiotemporal template for activation-inhibition of speech in ECoG
by: Easthope, Eric
Published: (2024)
by: Easthope, Eric
Published: (2024)
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
by: Sato, Motoshige, et al.
Published: (2026)
by: Sato, Motoshige, et al.
Published: (2026)
Learning spatial hearing via innate mechanisms
by: Chu, Yang, et al.
Published: (2020)
by: Chu, Yang, et al.
Published: (2020)
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
by: Geirnaert, Simon, et al.
Published: (2025)
by: Geirnaert, Simon, et al.
Published: (2025)
Brain-tuned Speech Models Better Reflect Speech Processing Stages in the Brain
by: Moussa, Omer, et al.
Published: (2025)
by: Moussa, Omer, et al.
Published: (2025)
Understanding Auditory Evoked Brain Signal via Physics-informed Embedding Network with Multi-Task Transformer
by: Ma, Wanli, et al.
Published: (2024)
by: Ma, Wanli, et al.
Published: (2024)
R&B -- Rhythm and Brain: Cross-subject Decoding of Music from Human Brain Activity
by: Ferrante, Matteo, et al.
Published: (2024)
by: Ferrante, Matteo, et al.
Published: (2024)
Human Brain Exhibits Distinct Patterns When Listening to Fake Versus Real Audio: Preliminary Evidence
by: Salehi, Mahsa, et al.
Published: (2024)
by: Salehi, Mahsa, et al.
Published: (2024)
Towards an End-to-End Framework for Invasive Brain Signal Decoding with Large Language Models
by: Feng, Sheng, et al.
Published: (2024)
by: Feng, Sheng, et al.
Published: (2024)
Mode-conditioned music learning and composition: a spiking neural network inspired by neuroscience and psychology
by: Liang, Qian, et al.
Published: (2024)
by: Liang, Qian, et al.
Published: (2024)
Brain2Music: Reconstructing Music from Human Brain Activity
by: Denk, Timo I., et al.
Published: (2023)
by: Denk, Timo I., et al.
Published: (2023)
Towards Dynamic Neural Communication and Speech Neuroprosthesis Based on Viseme Decoding
by: Park, Ji-Ha, et al.
Published: (2025)
by: Park, Ji-Ha, et al.
Published: (2025)
Scaling Law in Neural Data: Non-Invasive Speech Decoding with 175 Hours of EEG Data
by: Sato, Motoshige, et al.
Published: (2024)
by: Sato, Motoshige, et al.
Published: (2024)
Systematic review of self-supervised foundation models for brain network representation using electroencephalography
by: Portmann, Hannah, et al.
Published: (2026)
by: Portmann, Hannah, et al.
Published: (2026)
Predicting Artificial Neural Network Representations to Learn Recognition Model for Music Identification from Brain Recordings
by: Akama, Taketo, et al.
Published: (2024)
by: Akama, Taketo, et al.
Published: (2024)
The time course of visuo-semantic representations in the human brain is captured by combining vision and language models
by: Rong, Boyan, et al.
Published: (2025)
by: Rong, Boyan, et al.
Published: (2025)
Multi-modal brain encoding models for multi-modal stimuli
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
Decoding Selective Auditory Attention to Musical Elements in Ecologically Valid Music Listening
by: Akama, Taketo, et al.
Published: (2025)
by: Akama, Taketo, et al.
Published: (2025)
Condition-Invariant fMRI Decoding of Speech Intelligibility with Deep State Space Model
by: Sung, Ching-Chih, et al.
Published: (2025)
by: Sung, Ching-Chih, et al.
Published: (2025)
Using i-vectors for subject-independent cross-session EEG transfer learning
by: Lasko, Jonathan, et al.
Published: (2024)
by: Lasko, Jonathan, et al.
Published: (2024)
Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy
by: Kopar, Serli, et al.
Published: (2026)
by: Kopar, Serli, et al.
Published: (2026)
Towards auditory attention decoding with noise-tagging: A pilot study
by: Scheppink, H. A., et al.
Published: (2024)
by: Scheppink, H. A., et al.
Published: (2024)
Similar Items
-
Neural2Speech: A Transfer Learning Framework for Neural-Driven Speech Reconstruction
by: Li, Jiawei, et al.
Published: (2023) -
Speech language models lack important brain-relevant semantics
by: Oota, Subba Reddy, et al.
Published: (2023) -
A computational loudness model for electrical stimulation with cochlear implants
by: Alvarez, Franklin, et al.
Published: (2025) -
Prosody of speech production in latent post-stroke aphasia
by: Zhang, Cong, et al.
Published: (2024) -
Linguists should learn to love speech-based deep learning models
by: Kloots, Marianne de Heer, et al.
Published: (2025)