Phenotype Execution and Modelling Architecture (PhEMA) to support disease surveillance and real-world evidence studies: English sentinel network evaluation.
Jamie, G.; Elson, W.; de Lusignan, S.; Kar, D.; Wimalaratna, R.; Hoang, U.; Meza-Torres, B.; Forbes, A.; Hinton, W.; Anand, S.; Ferreira, F.; Ordonez-Mena, J.; Agrawal, U.; Byford, R.
Show abstract
ObjectiveTo evaluate Phenotype Execution and Modelling Architecture (PhEMA), to express sharable phenotypes using Clinical Query Language (CQL) and intensional SNOMED CT Fast Healthcare Interoperability Resources (FHIR) valuesets, for exemplar chronic disease, sociodemographic risk factor and surveillance phenotypes. MethodWe curated three phenotypes: Type 2 diabetes (T2DM), excessive alcohol use and incident influenza-like illness (ILI) using CQL to define clinical and administrative logic. We defined our phenotypes with valuesets, using SNOMEDs hierarchy and expression constraint language (ECL), and CQL, combining valuesets and adding temporal elements where needed. We compared the count of cases found using PhEMA with our existing approach using convenience datasets. ResultsThe T2DM phenotype could be defined as two intensionally defined SNOMED valuesets and a CQL script. It increased the prevalence from 7.2% to 7.3%. Excess alcohol phenotype was defined by valuesets that added qualitative clinical terms to the quantitative conceptual definitions we currently use; this change increased prevalence by 58%, from 1.2% to 1.9%. We created an ILI valueset with SNOMED concepts, adding a temporal element using CQL to differentiate new episodes. This increased the weekly incidence in our convenience sample (weeks 26 to 38) from 0.95 cases to 1.11 cases per 100,000 people. ConclusionsPhenotypes for surveillance and research can be described fully and comprehensibly using CQL and intensional FHIR valuesets. Our use case phenotypes identified a greater number of cases, whilst anticipated from excessive alcohol this was not for our other variable. This may have been due to our use of SNOMED CT hierarchy.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Transforming Estonian health data to the Observational Medical Outcomes Partnership (OMOP) Common Data Model: lessons learned 94%
- Determining prescriptions in electronic health care (EHR) data: methods for development of standardised, reproducible drug codelists 94%
- Trajectories: a framework for detecting temporal clinical event sequences from health data standardized to the OMOP Common Data Model 94%
Similar papers in this journal
- Replicating a COVID-19 study in a national England database to assess the generalisability of research with regional electronic health record data 96%
- The iDiabetes Platform: Enhanced Phenotyping of Patients with Diabetes for Precision Diagnosis, Prognosis and Treatment- study protocol for a cluster-randomised controlled study 94%
- Determining the feasibility of calculating pancreatic cancer risk scores for people with new-onset diabetes in primary care (DEFEND PRIME): study protocol 94%
Similar papers in this journal
- Development of a data-driven COVID-19 prognostication tool to inform triage and step-down care for hospitalised patients in Hong Kong: A population based cohort study 91%
- Development of a Mobile Application to Represent Food Intake in Inpatients: Dietary Data Systematization 91%
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 91%
Similar papers in this journal
Similar papers in this journal
- Clinical code sets and the problem of redundancy in code set repositories 94%
- CohortDiagnostics: phenotype evaluation across a network of observational data sources using population-level characterization 93%
- The Impact of Clinical Audits on Improving the Effectiveness of Type 2 Diabetes Mellitus (T2DM) CARE in Primary Health Centers. A Comprehensive Pre-post analysis through Multi-layered Intervention: The ICAE-DM CARE study protocol 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.