FLAMeS: A Robust Deep Learning Model for Automated Multiple Sclerosis Lesion Segmentation
Dereskewicz, E.; La Rosa, F.; dos Santos Silva, J.; Sizer, E.; Kohli, A.; Wynen, M.; Mullins, W. A.; Maggi, P.; Levy, S.; Onyemeh, K.; Ayci, B.; Solomon, A. J.; Assländer, J.; Al-Louzi, O.; Reich, D. S.; Sumowski, J. F.; Beck, E. S.
Show abstract
Background and PurposeAssessment of brain lesions on MRI is crucial for research in multiple sclerosis (MS). Manual segmentation is time consuming and inconsistent. We aimed to develop an automated MS lesion segmentation algorithm for T2-weighted fluid-attenuated inversion recovery (FLAIR) MRI. MethodsWe developed FLAIR Lesion Analysis in Multiple Sclerosis (FLAMeS), a deep learning-based MS lesion segmentation algorithm based on the nnU-Net 3D full-resolution U-Net and trained on 668 FLAIR 1.5 and 3 tesla scans from persons with MS. FLAMeS was evaluated on three external datasets: MSSEG-2 (n=14), MSLesSeg (n=51), and a clinical cohort (n=10), and compared to SAMSEG, LST-LPA, and LST-AI. Performance was assessed qualitatively by two blinded experts and quantitatively by comparing automated and ground truth lesion masks using standard segmentation metrics. ResultsIn a blinded qualitative review of 20 scans, both raters selected FLAMeS as the most accurate segmentation in 15 cases, with one rater favoring FLAMeS in two additional cases. Across all testing datasets, FLAMeS achieved a mean Dice score of 0.74, a true positive rate of 0.84, and an F1 score of 0.78, consistently outperforming the benchmark methods. For other metrics, including positive predictive value, relative volume difference, and false positive rate, FLAMeS performed similarly or better than benchmark methods. Most lesions missed by FLAMeS were smaller than 10 mm3, whereas the benchmark methods missed larger lesions in addition to smaller ones. ConclusionsFLAMeS is an accurate, robust method for MS lesion segmentation that outperforms other publicly available methods.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Fluid and White Matter Suppression Contrasts MRI Improves Deep Learning Detection of Multiple Sclerosis Cortical Lesions 97%
- LST-AI: a Deep Learning Ensemble for Accurate MS Lesion Segmentation 97%
- QSMRim-Net: Imbalance-Aware Learning for Identification of Chronic Active Multiple Sclerosis Lesions on Quantitative Susceptibility Maps 97%
Similar papers in this journal
- Lesion aware automated processing pipeline for multimodal neuroimaging stroke data and The Virtual Brain (TVB) 96%
- OpenMAP-T1: A Rapid Deep Learning Approach to Parcellate 280 Anatomical Regions to Cover the Whole Brain 96%
- Deep Bayesian networks for uncertainty estimation and adversarial resistance of white matter hyperintensity segmentation 95%
Similar papers in this journal
- Automated Generation of Cerebral Blood Flow Maps Using Deep Learning and Multiple Delay Arterial Spin-Labelled MRI 96%
- MASiVar: Multisite, Multiscanner, and Multisubject Acquisitions for Studying Variability in Diffusion Weighted Magnetic Resonance Imaging 94%
- Prospective Motion Correction and Automatic Segmentation of Penetrating Arteries in Phase Contrast MRI at 7 T 94%
Similar papers in this journal
- Deep neural networks allow expert-level brain meningioma detection, segmentation and improvement of current clinical practice 96%
- ANTsX: A dynamic ecosystem for quantitative biological and medical imaging 94%
- An AI-based segmentation and analysis pipeline for high-field MR monitoring of cerebral organoids 94%
Similar papers in this journal
- Multiple sclerosis cortical lesion detection with deep learning at ultra-high-field MRI 98%
- Impact of acquisition and modeling parameters on test-retest reproducibility of edited GABA+ 93%
- Hierarchical Bayesian Modelling Improves Microstructural Parameter Mapping in Diffusion and Exchange MRI Data 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.