Scaling laws for decoding images from brain activity

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Banville, Hubert, Benchetrit, Yohann, d'Ascoli, Stéphane, Rapin, Jérémy, King, Jean-Rémi
Format: Preprint
Publié: 2025
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866915125336735744
author Banville, Hubert
Benchetrit, Yohann
d'Ascoli, Stéphane
Rapin, Jérémy
King, Jean-Rémi
author_facet Banville, Hubert
Benchetrit, Yohann
d'Ascoli, Stéphane
Rapin, Jérémy
King, Jean-Rémi
contents Generative AI has recently propelled the decoding of images from brain activity. How do these approaches scale with the amount and type of neural recordings? Here, we systematically compare image decoding from four types of non-invasive devices: electroencephalography (EEG), magnetoencephalography (MEG), high-field functional Magnetic Resonance Imaging (3T fMRI) and ultra-high field (7T) fMRI. For this, we evaluate decoding models on the largest benchmark to date, encompassing 8 public datasets, 84 volunteers, 498 hours of brain recording and 2.3 million brain responses to natural images. Unlike previous work, we focus on single-trial decoding performance to simulate real-time settings. This systematic comparison reveals three main findings. First, the most precise neuroimaging devices tend to yield the best decoding performances, when the size of the training sets are similar. However, the gain enabled by deep learning - in comparison to linear models - is obtained with the noisiest devices. Second, we do not observe any plateau of decoding performance as the amount of training data increases. Rather, decoding performance scales log-linearly with the amount of brain recording. Third, this scaling law primarily depends on the amount of data per subject. However, little decoding gain is observed by increasing the number of subjects. Overall, these findings delineate the path most suitable to scale the decoding of images from non-invasive brain recordings.
format Preprint
id arxiv_https___arxiv_org_abs_2501_15322
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Scaling laws for decoding images from brain activity
Banville, Hubert
Benchetrit, Yohann
d'Ascoli, Stéphane
Rapin, Jérémy
King, Jean-Rémi
Image and Video Processing
Artificial Intelligence
Machine Learning
Neurons and Cognition
Generative AI has recently propelled the decoding of images from brain activity. How do these approaches scale with the amount and type of neural recordings? Here, we systematically compare image decoding from four types of non-invasive devices: electroencephalography (EEG), magnetoencephalography (MEG), high-field functional Magnetic Resonance Imaging (3T fMRI) and ultra-high field (7T) fMRI. For this, we evaluate decoding models on the largest benchmark to date, encompassing 8 public datasets, 84 volunteers, 498 hours of brain recording and 2.3 million brain responses to natural images. Unlike previous work, we focus on single-trial decoding performance to simulate real-time settings. This systematic comparison reveals three main findings. First, the most precise neuroimaging devices tend to yield the best decoding performances, when the size of the training sets are similar. However, the gain enabled by deep learning - in comparison to linear models - is obtained with the noisiest devices. Second, we do not observe any plateau of decoding performance as the amount of training data increases. Rather, decoding performance scales log-linearly with the amount of brain recording. Third, this scaling law primarily depends on the amount of data per subject. However, little decoding gain is observed by increasing the number of subjects. Overall, these findings delineate the path most suitable to scale the decoding of images from non-invasive brain recordings.
title Scaling laws for decoding images from brain activity
topic Image and Video Processing
Artificial Intelligence
Machine Learning
Neurons and Cognition
url https://arxiv.org/abs/2501.15322