Back

Evaluating the Concordance between ICD-10 and Stroke Severity as Measured by the NIHSS

Taha, M.; Habib, M.; Lomachinsky, V.; Hadar, P.; Newhouse, J. P.; Schwamm, L. H.; Blacker, D.; Moura, L. M. V. R.

2024-02-23 health policy
10.1101/2024.02.21.24303177 medRxiv
Show abstract

BackgroundThe National Institutes of Health Stroke Scale (NIHSS) scores have been used to evaluate Acute Ischemic Stroke (AIS) severity in clinical settings. Through the International Classification of Diseases, Tenth Revision Code (ICD-10), documentation of NIHSS scores has been made possible for administrative purposes and has since been increasingly adopted in insurance claims. Per CMS guidelines, the stroke ICD-10 diagnosis code must be documented by the treating physician, but ICD-10 NIHSS scores can be documented by any healthcare provider involved in the patients care. Accuracy of the administratively collected NIHSS compared to expert clinical evaluation as documented in the Paul Coverdell registry is however still uncertain. MethodsLeveraging a linked dataset comprised of the Paul Coverdell National Acute Stroke Program (PCNASP) clinical registry and probabilistically matched individuals on Medicare Claims data, we sampled patients aged 65 and above admitted for AIS across nine states, from 2016 to 2019. We excluded those lacking documentation for either clinical or ICD-10 based NIHSS scores. We then examined score concordance from both databases and measured discordance as the absolute difference between the PCNASP and ICD-10-based NIHSS scores. ResultsAmong 66,837 matched patients, mean NIHSS scores for PCNASP and Medicare ICD-10 were 7.26 (95% CI: 7.20 - 7.32) and 7.40 (95% CI: 7.34 - 7.46), respectively. Concordance between the two scores was high as indicated by an intraclass correlation coefficient of 0.93. ConclusionThe high concordance between clinical and ICD-10 NIHSS scores highlights the latters potential as measure of stroke severity derived from structured claims data.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.