Automatic detection of abnormal clinical EEG: comparison of a finetuned foundation model with two deep learning models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bussalb, Aurore, Gac, François Le, Jubien, Guillaume, Rahmouni, Mohamed, Bettinardi, Ruggero G., de Oliveira, Pedro Marinho R., Derambure, Phillipe, Gaspard, Nicolas, Jonas, Jacques, Maillard, Louis, Vercueil, Laurent, Vespignani, Hervé, Laval, Philippe, Koessler, Laurent, Gimenez, Ulysse
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912398771748864
author Bussalb, Aurore
Gac, François Le
Jubien, Guillaume
Rahmouni, Mohamed
Bettinardi, Ruggero G.
de Oliveira, Pedro Marinho R.
Derambure, Phillipe
Gaspard, Nicolas
Jonas, Jacques
Maillard, Louis
Vercueil, Laurent
Vespignani, Hervé
Laval, Philippe
Koessler, Laurent
Gimenez, Ulysse
author_facet Bussalb, Aurore
Gac, François Le
Jubien, Guillaume
Rahmouni, Mohamed
Bettinardi, Ruggero G.
de Oliveira, Pedro Marinho R.
Derambure, Phillipe
Gaspard, Nicolas
Jonas, Jacques
Maillard, Louis
Vercueil, Laurent
Vespignani, Hervé
Laval, Philippe
Koessler, Laurent
Gimenez, Ulysse
contents Electroencephalography (EEG) is commonly used by physicians for the diagnosis of numerous neurological disorders. Due to the large volume of EEGs requiring interpretation and the specific expertise involved, artificial intelligence-based tools are being developed to assist in their visual analysis. In this paper, we compare two deep learning models (CNN-LSTM and Transformer-based) with BioSerenity-E1, a recently proposed foundation model, in the task of classifying entire EEG recordings as normal or abnormal. The three models were trained or finetuned on 2,500 EEG recordings and their performances were evaluated on two private and one public datasets: a large multicenter dataset annotated by a single specialist (dataset A composed of n = 4,480 recordings), a small multicenter dataset annotated by three specialists (dataset B, n = 198), and the Temple University Abnormal (TUAB) EEG corpus evaluation dataset (n = 276). On dataset A, the three models achieved at least 86% balanced accuracy, with BioSerenity-E1 finetuned achieving the highest balanced accuracy (89.19% [88.36-90.41]). BioSerenity-E1 finetuned also achieved the best performance on dataset B, with 94.63% [92.32-98.12] balanced accuracy. The models were then validated on TUAB evaluation dataset, whose corresponding training set was not used during training, where they achieved at least 76% accuracy. Specifically, BioSerenity-E1 finetuned outperformed the other two models, reaching an accuracy of 82.25% [78.27-87.48]. Our results highlight the usefulness of leveraging pre-trained models for automatic EEG classification: enabling robust and efficient interpretation of EEG data with fewer resources and broader applicability.
format Preprint
id arxiv_https___arxiv_org_abs_2505_21507
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Automatic detection of abnormal clinical EEG: comparison of a finetuned foundation model with two deep learning models
Bussalb, Aurore
Gac, François Le
Jubien, Guillaume
Rahmouni, Mohamed
Bettinardi, Ruggero G.
de Oliveira, Pedro Marinho R.
Derambure, Phillipe
Gaspard, Nicolas
Jonas, Jacques
Maillard, Louis
Vercueil, Laurent
Vespignani, Hervé
Laval, Philippe
Koessler, Laurent
Gimenez, Ulysse
Neurons and Cognition
Machine Learning
Signal Processing
Electroencephalography (EEG) is commonly used by physicians for the diagnosis of numerous neurological disorders. Due to the large volume of EEGs requiring interpretation and the specific expertise involved, artificial intelligence-based tools are being developed to assist in their visual analysis. In this paper, we compare two deep learning models (CNN-LSTM and Transformer-based) with BioSerenity-E1, a recently proposed foundation model, in the task of classifying entire EEG recordings as normal or abnormal. The three models were trained or finetuned on 2,500 EEG recordings and their performances were evaluated on two private and one public datasets: a large multicenter dataset annotated by a single specialist (dataset A composed of n = 4,480 recordings), a small multicenter dataset annotated by three specialists (dataset B, n = 198), and the Temple University Abnormal (TUAB) EEG corpus evaluation dataset (n = 276). On dataset A, the three models achieved at least 86% balanced accuracy, with BioSerenity-E1 finetuned achieving the highest balanced accuracy (89.19% [88.36-90.41]). BioSerenity-E1 finetuned also achieved the best performance on dataset B, with 94.63% [92.32-98.12] balanced accuracy. The models were then validated on TUAB evaluation dataset, whose corresponding training set was not used during training, where they achieved at least 76% accuracy. Specifically, BioSerenity-E1 finetuned outperformed the other two models, reaching an accuracy of 82.25% [78.27-87.48]. Our results highlight the usefulness of leveraging pre-trained models for automatic EEG classification: enabling robust and efficient interpretation of EEG data with fewer resources and broader applicability.
title Automatic detection of abnormal clinical EEG: comparison of a finetuned foundation model with two deep learning models
topic Neurons and Cognition
Machine Learning
Signal Processing
url https://arxiv.org/abs/2505.21507