Built Environment Features Obtained from Google Street View Are Associated with Coronary Artery Disease Prevalence: A Deep-Learning Framework
Chen, Z.; Khalifa, Y.; Dazard, J.-E.; Motairek, I.; Sadeer, A.-K. G.; Rajagopalan, S.
Show abstract
BackgroundBuilt environment plays an important role in development of cardiovascular disease. Tools to evaluate the built environment using machine vision and informatic approaches has been limited. We sought to investigate the association between machine vision-based built environment and prevalence of cardiometabolic disease in urban cities. MethodsThis cross-sectional study used features extracted from Google Street view (GSV) images to measure the built environment and link them with prevalence of cardiometabolic disease. Convolutional neural networks, light gradient boosting machines and activation maps were utilized to predict health outcomes and identify feature associations with coronary heart disease (CHD). The study obtained 0.53 million GSV images covering 789 census tracts in 7 cities (Cleveland, OH; Fremont, CA; Kansas City, MO; Detroit, MI; Bellevue, WA; Brownsville, TX; and Denver, CO). Analyses were conducted from February 2022 to December 2022. We used census tract-level data from the Centers for Disease Control and Preventions PLACES dataset. Main outcomes included census tract-level estimated prevalence of CHD based on GSV built environment features. ResultsBuilt environment features extracted from GSV using deep learning predicted 63% of the census tract variation in CHD prevalence. The ExtraTrees Regressor achieved the best result among all models with the lowest average mean absolute error of 1.11% and Root mean square of error of 1.58. The addition of GSV features outperformed and improved a model that only included census-tract level age, sex, race, income and education. Activation maps from the features revealed a set of neighborhood features represented by buildings and roads associated with CHD prevalence. ConclusionsIn this cross-sectional study, a significant portion of CHD prevalence were explained by GSV-based built environment factors analyzed using deep learning, independent of census tract demographics. Machine vision enabled assessment of the built environment could help play a significant role in designing and improving heart-heathy cities.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Enhanced machine learning and hybrid ensemble approaches for coronary heart disease prediction 92%
- Rural Roads to Cognitive Resilience (RRR): A prospective cohort study protocol 91%
- A machine learning approach to identifying important features for achieving step thresholds in individuals with chronic stroke 91%
Similar papers in this journal
Similar papers in this journal
- Opportunistic Assessment of Ischemic Heart Disease Risk Using Abdominopelvic Computed Tomography and Medical Record Data: a Multimodal Explainable Artificial Intelligence Approach 93%
- Can machine learning improve risk prediction of incident hypertension? An internal method comparison and external validation of the Framingham risk model using HUNT Study data 91%
- High-Dimensional Multinomial Multiclass Severity Scoring of COVID-19 Pneumonia Using CT Radiomics Features and Machine Learning Algorithms 90%
Similar papers in this journal
- Using ECG Machine Learning for Detection of Cardiovascular Disease in African American Men and Women: the Jackson Heart Study 91%
- Prediction of Cardiovascular Markers and Diseases Using Retinal Fundus Images and Deep Learning: A Systematic Scoping Review 91%
- Explainable AI in Deep Learning-based Detection of Aortic Elongation on Chest X-ray Images 89%
Similar papers in this journal
- Automated Image Transcription for Perinatal Blood Pressure Monitoring Using Mobile Health Technology 90%
- Designing a computer-assisted diagnosis system for cardiomegaly detection and radiology report generation 90%
- Population Analysis Of Mortality Risk: Predictive Models Using Motion Sensors For 100,000 Participants In The UK Biobank National Cohort 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.