Back

Some statistical theory for interpreting reference distributions

Alpay, B. A.; Higgins, J. M.; Desai, M. M.

2024-07-24 pathology
10.1101/2024.07.23.24309680 medRxiv
Show abstract

Reference distributions quantify the extremeness of clinical test results, typically relative to those of a healthy population. Intervals of these distributions are used in medical decision-making, but while there is much guidance for constructing them, the statistics of interpreting them for diagnosis have been less explored. Here we work directly in terms of the reference distribution, defining it as the likelihood in a posterior calculation of the probability of disease. We thereby identify assumptions of the conventional interpretation of reference distributions, criteria for combining tests, and considerations for personalizing interpretation of results from reference data. Theoretical reasoning supports that non-healthy variation be taken into account when possible, and that combining and personalizing tests call for careful statistical modeling.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.