Cross-Institutional European Evaluation and Validation of Automated Multilabel Segmentation for Acute Intracerebral Hemorrhage and Complications
Nawabi, J.; Baumgaertner, G. L.; Schulze-Weddige, S.; Dell'Orco, A.; Morotti, A.; Mazzacane, F.; Kniep, H. C.; Schlunk, F.; Boehmer, M. F.; Akkurt, B.; Orth, T.; Weissflog, J.-S.; Schumann, M.; Sporns, P.; Scheel, M.; Hanning, U.; Fiehler, J.; Penzkofer, T.
Show abstract
PurposeTo evaluate a nnU-Net-based deep learning for automated segmentation of intracerebral hemorrhage (ICH), intraventricular hemorrhage (IVH), and perihematomal edema (PHE) on noncontrast CT scans. Materials and MethodsRetrospective data from acute ICH patients admitted at four European stroke centers (2017-2019), along healthy controls (2022-2023), were analyzed. nnU-Net was trained (n=775) using a 5-fold cross-valiadtion approach, tested (n=189), and seperatly validated on internal (n=121), external (n=169), and diverse ICH etiologies (n=175) datasets. Interrater-validated ground truth served as the reference standard. Lesion detection, segmentation, and volumetric accuracy were measured, alongside time efficiency versus manual segmentation. ResultsTest set results revealed high nnU-Net accuracy (median Dice Similartiy Coefficient (DSC): ICH 0.91, IVH 0.76, PHE 0.71) and volumetric correlation (ICH, IVH: r=0.99; PHE: r=0.92). Sensitivities were high (ICH, PHE: 99%; IVH: 97%), with IVH detection specificities and sensitivities >90% for volumes up to 0.2 ml. Anatomical-specific metrics showed higher performance for lobar and deep hemorrhages (median DSC 0.90 and 0.92, respectively) and lower for brainstem (median DSC 0.70). Concurrent hemorrhages did not affect accuracy, p> 0.05. Across validation sets, segmentation precision was consistent, especially for ICH (median DSC 0.85-0.90), with PHE slightly lower (median DSC 0.61-0.66) and IVH best in the second and third set (median DSC 0.80). Average processing time was 18.2 seconds versus 18.01 minutes manually. ConclusionThe nnU-Net provides reliable, time-efficient ICH, IVH, and PHE segmentation, validated across various clinical settings, with excellent anatomical-specific performance for lobar and deep hemorrhages. It shows promise for enhancing clinical workflow and research initiatives.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Deep neural networks allow expert-level brain meningioma detection, segmentation and improvement of current clinical practice 95%
- Does contrast-enhancement improve visualisation of lenticulostriate arteries in cerebral small vessel disease using time-of-flight magnetic resonance angiography at 7 Tesla? 92%
- An AI-based segmentation and analysis pipeline for high-field MR monitoring of cerebral organoids 92%
Similar papers in this journal
- Deep-learning-based white matter lesion volume in CT is associated with outcome after acute ischemic stroke 95%
- Evaluating Large Language Model-Generated Brain MRI Protocols: Performance of GPT4o, o3-mini, DeepSeek-R1 and Qwen2.5-72B 93%
- From Community Acquired Pneumonia to COVID-19: A Deep Learning Based Method for Quantitative Analysis of COVID-19 on thick-section CT Scans 90%
Similar papers in this journal
- Scaling behaviors of deep learning and linear algorithms for the prediction of stroke severity 95%
- Machine learning-based prediction of motor status in glioma patients using diffusion MRI metrics along the corticospinal tract 94%
- Deep Learning disconnectomes to accelerate and improve long-term predictions for post-stroke symptoms 92%
Similar papers in this journal
- Fluid and White Matter Suppression Contrasts MRI Improves Deep Learning Detection of Multiple Sclerosis Cortical Lesions 93%
- QSMRim-Net: Imbalance-Aware Learning for Identification of Chronic Active Multiple Sclerosis Lesions on Quantitative Susceptibility Maps 92%
- Right hemispheric white matter hyperintensities improve the prediction of spatial neglect severity in acute stroke 91%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.