DETECT: Feature extraction method for disease trajectory modeling
Singhal, P.; Guare, L.; Morse, C.; Byrska-Bishop, M.; Guerraty, M. A.; Kim, D.; Ritchie, M. D.; Verma, A.
Show abstract
Modeling with longitudinal electronic health record (EHR) data proves challenging given the high dimensionality, redundancy, and noise captured in EHR. In order to improve precision medicine strategies and identify predictors of disease risk in advance, evaluating meaningful patient disease trajectories is essential. In this study, we develop the algorithm DiseasE Trajectory fEature extraCTion (DETECT) for feature extraction and trajectory generation in high-throughput temporal EHR data. This algorithm can 1) simulate longitudinal individual-level EHR data, specified to user parameters of scale, complexity, and noise and 2) use a convergent relative risk framework to test intermediate codes occurring between a specified index code(s) and outcome code(s) to determine if they are predictive features of the outcome. We benchmarked our method on simulated data and generated real-world disease trajectories using DETECT in a cohort of 145,575 individuals diagnosed with hypertension in Penn Medicine EHR for severe cardiometabolic outcomes.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Using indication embeddings to represent patient health for drug safety studies 94%
- Modeling physician variability to prioritize relevant medical record information 94%
- Trajectories: a framework for detecting temporal clinical event sequences from health data standardized to the OMOP Common Data Model 94%
Similar papers in this journal
- Mining for Equitable Health: Assessing the Impact of Missing Data in Electronic Health Records 96%
- ARCH: Large-scale Knowledge Graph via Aggregated Narrative Codified Health Records Analysis 94%
- A methodology of phenotyping ICU patients from EHR data: high-fidelity, personalized, and interpretable phenotypes estimation 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.