Ensemble learning for robust knee cartilage segmentation: data from the osteoarthritis initiative.
Peake, E.; Chevasson, R.; Pszczolkowski, S.; Auer, D. P.; Arthofer, C.
Show abstract
PurposeTo evaluate the performance of an ensemble learning approach for fully automated cartilage segmentation on knee magnetic resonance images of patients with osteoarthritis. Materials and MethodsThis retrospective study of 88 participants with knee osteoarthritis involved the study of three-dimensional (3D) double echo steady state (DESS) MR imaging volumes with manual segmentations for 6 different compartments of cartilage (Data available from the Osteoarthritis Initiative). We propose ensemble learning to boost the sensitivity of our deep learning method by combining predictions from two models, a U-Net for the segmentation of two labels (cartilage vs background) and a multi-label U-Net for specific cartilage compartments. Segmentation accuracy is evaluated using Dice coefficient, while volumetric measures and Bland Altman plots provide complimentary information when assessing segmentation results. ResultsOur model showed excellent accuracy for all 6 cartilage locations: femoral 0.88, medial tibial 0.84, lateral tibial 0.88, patellar 0.85, medial meniscal 0.85 and lateral meniscal 0.90. The average volume correlation was 0.988, overestimating volume by 9% {+/-} 14% over all compartments. Simple post processing creates a single 3D connected component per compartment resulting in higher anatomical face validity. ConclusionOur model produces automated segmentation with high Dice coefficients when compared to expert manual annotations and leads to the recovery of missing labels in the manual annotations, while also creating smoother, more realistic boundaries avoiding slice discontinuity artifacts present in the manual annotations. Key ResultsO_LICombining a 2-label U-Net (cartilage vs background) with a multi-class U-Net for segmentation of cartilage compartment boosts the accuracy of our deep learning model leading to the recovery of missing annotations in the manual dataset. C_LIO_LIAutomatically generated segmentations have high Dice coefficients (0.85-0.90) and reduce inter-slice discontinuity artefact caused by slice wise delineation. C_LIO_LIModel refinement yields more anatomically plausible segmentations where each cartilage label is composed of only a single 3D region of interest. C_LI
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- pyKNEEr: An image analysis workflow for open and reproducible research on femoral knee cartilage 96%
- A registration strategy to characterize DTI-observed changes in skeletal muscle architecture due to passive shortening 95%
- Real-time three-dimensional MRI for the assessment of dynamic carpal instability 93%
Similar papers in this journal
- Liver Volumetry from Magnetic Resonance Images with Convolutional Neural Networks 93%
- Simulated Diagnostic Performance of Ultra-Low-Field MRI: Harnessing Open-Access Datasets to Evaluate Novel Devices 93%
- Increased Brain Volumetric Measurement Precision from Multi-Site 3D T1-weighted 3T Magnetic Resonance Imaging by Correcting Geometric Distortions 92%
Similar papers in this journal
- MRI-derived Articular Cartilage Strains Predict Patient-Reported Outcomes Six Months Post Anterior Cruciate Ligament Reconstruction 94%
- Correcting B0 inhomogeneity-induced distortions in whole-body diffusion MRI of bone metastases 94%
- An AI-based segmentation and analysis pipeline for high-field MR monitoring of cerebral organoids 93%
Similar papers in this journal
- Multi-frame biomechanical and relaxometry analysis during in vivo loading of the human knee by spiral dualMRI and compressed sensing 95%
- Image- vs. histogram-based considerations in semantic segmentation of pulmonary hyperpolarized gas images 93%
- Prospective Motion Correction and Automatic Segmentation of Penetrating Arteries in Phase Contrast MRI at 7 T 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.