Asyndromic Surveillance of New York City Emergency Department Diagnoses with the Tree-Temporal Scan Statistic
Greene, S. K.; Levin-Rector, A.; Kulldorff, M.; Lall, R.
Show abstract
ObjectivesIllness trends are typically monitored by reportable disease and syndromic surveillance systems, but unanticipated health issues might not be captured. Using diagnosis codes, the New York City Health Department developed a data mining method to detect unusual increases in emergency department (ED) visits for any reason. MethodsWe applied the tree-temporal scan statistic in TreeScan software to ICD-10-CM diagnosis codes for ED visits. We searched for unusual citywide increases in ED visits or hospital admissions, over any recent time period, and at any part of and level on the ICD-10-CM tree. We conducted proof-of-concept analyses for March 2020 when COVID-19 emerged, then investigated signals detected in daily, automated analyses during April-August 2025. ResultsIf TreeScan analyses had been in place, then increasing hospital admissions for viral pneumonia (J12) would have triggered a signal on March 13, 2020, two days before widespread COVID-19 community transmission was announced. An extreme heat event in June 2025 triggered a signal for admissions for acute kidney failure (N17), prompting outreach to dialysis networks. A sustained signal for hand, foot, and mouth disease (B08.4) prompted outreach to child care programs. Other signals supported situational awareness, including a seasonal increase for swimmers ear (H60.33) and burns (T30.0) related to consumer fireworks. Practice ImplicationsTreeScan quickly detected credible increases in various diagnoses without pre-specification, from minor to severe, rare to common, acute to sustained, and foreseen to unforeseen. TreeScan can strengthen surveillance for health issues related to new pathogens, non-notifiable conditions, environmental exposures, and mass gatherings.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Finding Long-COVID: Temporal Topic Modeling of Electronic Health Records from the N3C and RECOVER Programs 94%
- Novel clinical subphenotypes in COVID-19: derivation, validation, prediction, temporal patterns, and interaction with social determinants of health 93%
- Identifying clusters of people with Multiple Long-Term Conditions using Large Language Models: a population-based study 92%
Similar papers in this journal
- Clinical Outcomes, Costs, and Cost-effectiveness of Strategies for People Experiencing Sheltered Homelessness During the COVID-19 Pandemic 93%
- Children with SARS-CoV-2 in the National COVID Cohort Collaborative (N3C) 92%
- COVID-19 cases and hospitalizations averted by case investigation and contact tracing in the United States 92%
Similar papers in this journal
- Clinical prediction rule for SARS-CoV-2 infection from 116 U.S. emergency departments 94%
- Reduced turnaround times through multi-sectoral collaboration during the first surge of SARS-CoV-2 in Louisiana, March-April 2020 93%
- A machine learning-based phenotype for long COVID in children: an EHR-based study from the RECOVER program 93%
Similar papers in this journal
- Risk factors for severe COVID-19 in hospitalized children in Canada: A national prospective study from March 2020–May 2021 91%
- Vaccination reduces need for emergency care in breakthrough COVID-19 infections: A multicenter cohort study 91%
- Phylogenetic estimates of SARS-CoV-2 introductions into Washington State 90%
Similar papers in this journal
- Risk Factor Analysis for Extended-Spectrum Beta-Lactamase Producing Enterobacterales Colonization or Infection: Evaluation of a Novel Approach to Assess Local Prevalence as a Risk Factor 91%
- N95 Filtering Facepiece Respirators Remain Effective After Extensive Reuse During the COVID-19 Pandemic 90%
- Routine saliva testing for the identification of silent COVID-19 infections in healthcare workers 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.