Back

Bias in (sero)prevalence estimates

Haile, S. R.

2022-11-29 epidemiology
10.1101/2022.11.24.22282720 medRxiv
Show abstract

BackgroundThe COVID-19 pandemic has led to many studies of seroprevalence. A number of methods exist in the statistical literature to correctly estimate disease prevalence or seroprevalence in the presence of diagnostic test misclassification, but these methods seem to be less known and not routinely used in the public health literature. We aimed to examine how widespread the problem is in recent publications, and to quantify the magnitude of bias introduced when correct methods are not used. MethodsA systematic review was performed to estimate how often public health researchers accounted for diagnostic test performance in estimates of seroprevalence. Using straightforward calculations, we estimated the amount of bias introduced when reporting the proportion of positive test results instead of using sensitivity and specificity to estimate disease prevalence. ResultsOf the seroprevalence studies sampled, 78% (95% CI 72% to 82%) failed to account for sensitivity and specificity. Expected bias is often more than is desired in practice, ranging from 1% to 12%. ConclusionsResearchers conducting studies of prevalence should correctly account for test sensitivity and specificity in their statistical analysis.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.