CViT-ESP: Lightweight Pre-trained Vision Transformers for EEG-based Epileptic Seizure Prediction
Mohammad, U.; Parani, P.; Saeed, F.
Show abstract
Background and Objective Epileptic seizure prediction is a critical challenge requiring the discrimination of subtle preictal physiological changes from interictal brain activity. While deep learning has shown promise in this domain, existing models often face limitations due to small EEG datasets, high computational costs for training from scratch, and a lack of patient-independent generalizability. In this paper, we present a novel framework for EEG-based seizure prediction that leverages pre-trained Vision Transformers (ViTs) through custom architectural modifications and optimized re-training strategies. Methods Our primary contributions include: [bullet]CVIT-ESP: A family of vision transformer architectures that replaces standard patch embedding layers with custom N-dimensional CNN stages to refine EEG representations. [bullet] ESPFormer: A lightweight, custom-designed transformer specifically engineered to mitigate overfitting on limited-scale EEG datasets. We identified optimal fine-tuning combinations for transformer blocks by devising a heuristic search-space reduction strategy, significantly reducing the training complexity. We validated our methods using the patient-independent MLSPred-Bench, involving 12 diverse benchmarks with varying seizure prediction horizons. Results Results demonstrate a clear progression in performance: while prior ResNet and vanilla Transformer models achieved an AUC-ROC of 69.0%, our CVIT-ESP architectures achieved the highest performance with a maximum average AUC of 76.4%. Conclusions These findings suggest that adapting pre-trained ViTs with domain-specific CNN front-ends and strategic fine-tuning offers a robust, generalizable, and resource-efficient path forward for clinical seizure prediction systems. Our code is available at: https://github.com/pcdslab/CVitEsp and https://github.com/pcdslab/ESPFormer
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Multiscale predictive modeling robustly improves the accuracy of pseudo-prospective seizure forecasting in drug-resistant epilepsy 95%
- PyHFO: Lightweight Deep Learning-poweredEnd-to-End High-Frequency Oscillations AnalysisApplication 94%
- Interpretable EEG Biomarkers for Neurological Disease Models in Mice Using Bag-of-Waves Classifiers 94%
Similar papers in this journal
- Virtual epilepsy patient cohort: generation and evaluation 93%
- Data-driven method to infer the seizure propagation patterns in an epileptic brain from intracranial electroencephalography 93%
- Evidence for spreading seizure as a cause of theta-alpha activity electrographic pattern in stereo-EEG seizure recordings 91%
Similar papers in this journal
- Event Driven Neural Network on a Mixed Signal Neuromorphic Processor for EEG Based Epileptic Seizure Detection 94%
- NLP-based tools for localization of the Epileptogenic Zone in patients with drug-resistant focal epilepsy 94%
- Recurrent Neural Network-based Acute Concussion Classifier using Raw Resting State EEG Data 93%
Similar papers in this journal
- Intracortical neural activity distal to seizure-onset-areas predicts human focal seizures 92%
- Wavelet Phase Coherence of Ictal Scalp EEG-Extracted Muscle Activity (SMA) as a Biomarker for Sudden Unexpected Death in Epilepsy (SUDEP) 92%
- Automatic diagnostics of electroencephalography pathology based on multi-domain feature fusion 92%
Similar papers in this journal
- Noninvasive, automated and reliable detection of spreading depolarizations in severe traumatic brain injury using scalp EEG 93%
- Respiratory modulations of cortical excitabilityand interictal spike timing in focal epilepsy - a case report 90%
- Non-vectorial Integration of Intersectional Short-Pulse Stimulation Enables Enhanced Deep Brain Modulation and Effective Seizure Control 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.