Machine learning to classify left ventricular hypertrophy using ECG feature extraction by variational autoencoder
Gupta, A.; Harvey, C. J.; DeBauge, A.; Shomaji, S.; Yao, Z.; Noheria, A.
Show abstract
BackgroundTraditional ECG criteria for left ventricular hypertrophy (LVH) have modest diagnostic yield. ObjectiveDevelop and validate machine learning models for LVH diagnosis from ECG. MethodsECG summary features (rate, intervals, axis), R-wave, S-wave and overall-QRS amplitudes, and QRS voltage-time integrals (VTIQRS) were extracted from 12-lead, vectorcardiographic X-Y-Z-lead, and 3D (L2 norm) representative-beat ECGs. Latent features (30 per ECG) were extracted using a variational autoencoder (trained on unselected >1 million ECGs) from X-Y-Z-lead representative-beat ECG signals. Logistic regression, random forest, light gradient boosted machine (LGBM), residual network (ResNet) and multilayer perceptron network (MLP) models using ECG features and sex, and a convolutional neural network (CNN) using ECG signals alone, were trained to predict LVH (left ventricular mass indexed in women >95 g/m2, men >115 g/m2) on 482,734 adult ECG-echocardiogram (within 45 days) pairs. ROC-AUCs for LVH classification are reported from a separate hold-out test set. ResultsIn the test set (n=54,984), AUC for LVH classification was higher for ML models using ECG features (LGBM 0.794, MLP 0.793, ResNet 0.795) compared with the best individual ECG variable (VTIQRS-Z 0.707), the best traditional criterion (Cornell voltage-duration product 0.716), and the CNN using ECG signals (0.788). Among patients without LVH who had a follow-up echocardiogram >1 (closest to 5) year later, LGBM false positives, compared to true negatives, had a 3.07 (95% CI 2.44, 3.86)-fold higher odds of developing future LVH (p<0.0001). ConclusionsML models are superior to traditional ECG criteria to classify LVH. Models trained on extracted ECG features, including latent variational autoencoder representations, can outperform CNN models directly trained on ECG signals.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Relationship of mild to moderate impairment of left ventricular ejection fraction with fatal ventricular arrhythmic events in cardiac sarcoidosis 95%
- Smartwatch Facilitated Remote Health Care for Patients Undergoing Transcatheter Aortic Valve Replacement Amid COVID-19 Pandemic 95%
- Non-Invasive Scale Measurement of Cardiac Output Compared with the Gold-Standard Direct Fick Method: A Feasibility Study 95%
Similar papers in this journal
- Simple Models Versus Deep Learning in Detecting Low Ejection Fraction From The Electrocardiogram 96%
- Using ECG Machine Learning for Detection of Cardiovascular Disease in African American Men and Women: the Jackson Heart Study 96%
- Automated Echocardiographic Detection of Mitral Valve Prolapse and Mitral Regurgitation with Video-based Artificial Intelligence Algorithms 95%
Similar papers in this journal
- Predictors and outcomes of Cardiac Dyssynchrony among patients with heart failure attending Benjamin Mkapa Hospital in Dodoma, central Tanzania: A protocol of prospective-longitudinal study 95%
- T vector velocity: A new ECG biomarker for identifying drug effects on cardiac ventricular repolarization 95%
- {-}CardiOvascular examination in awake Orangutans (Pongo pygmaeus pygmaeus): Low-stress Echocardiography including Speckle Tracking imaging (the COOLEST method) 94%
Similar papers in this journal
- Prognostic value of compact myocardial thinning in patients with left ventricular non-compaction 96%
- Investigating Electrocardiographic Abnormalities in Patients with Coronary Microvascular Dysfunction 96%
- A Multicenter Evaluation of the Impact of Procedural and Pharmacological Interventions on Deep Learning-based Electrocardiographic Markers of Hypertrophic Cardiomyopathy 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.