Optimal two-phase sampling design for comparing accuracies of two binary classification rules

In this paper, we consider the design for comparing the performance of two binary classification rules, for example, two record linkage algorithms or two screening tests. Statistical methods are well developed for comparing these accuracy measures when the gold standard is available for every unit i...

Full description

Saved in:

Bibliographic Details
Published in	Statistics in medicine Vol. 33; no. 3; pp. 500 - 513
Main Authors	Xu, Huiping, Hui, Siu L., Grannis, Shaun
Format	Journal Article
Language	English
Published	England Blackwell Publishing Ltd 10.02.2014 Wiley Subscription Services, Inc
Subjects	Algorithms Biometry - methods Classification - methods diagnostic accuracy diagnostic test Female Humans Male Medical statistics Models, Statistical Parameter estimation positive predicted value Predictive Value of Tests record linkage Sampling Scheduling algorithms sensitivity specificity stratified sampling diagnostic accuracy positive predicted value record linkage specificity sensitivity stratified sampling diagnostic test
Online Access	Get full text

Cover

Loading…

More Information
Summary:	In this paper, we consider the design for comparing the performance of two binary classification rules, for example, two record linkage algorithms or two screening tests. Statistical methods are well developed for comparing these accuracy measures when the gold standard is available for every unit in the sample, or in a two‐phase study when the gold standard is ascertained only in the second phase in a subsample using a fixed sampling scheme. However, these methods do not attempt to optimize the sampling scheme to minimize the variance of the estimators of interest. In comparing the performance of two classification rules, the parameters of primary interest are the difference in sensitivities, specificities, and positive predictive values. We derived the analytic variance formulas for these parameter estimates and used them to obtain the optimal sampling design. The efficiency of the optimal sampling design is evaluated through an empirical investigation that compares the optimal sampling with simple random sampling and with proportional allocation. Results of the empirical study show that the optimal sampling design is similar for estimating the difference in sensitivities and in specificities, and both achieve a substantial amount of variance reduction with an over‐sample of subjects with discordant results and under‐sample of subjects with concordant results. A heuristic rule is recommended when there is no prior knowledge of individual sensitivities and specificities, or the prevalence of the true positive findings in the study population. The optimal sampling is applied to a real‐world example in record linkage to evaluate the difference in classification accuracy of two matching algorithms. Copyright © 2013 John Wiley & Sons, Ltd.
Bibliography:	Agency for Healthcare Research and Quality - No. R01HS018553 istex:827C87099059AA70B975ADD9460CCA1412DDE3F9 ArticleID:SIM5946 ark:/67375/WNG-F6CKVV60-N SourceType-Scholarly Journals-1 ObjectType-Feature-1 content type line 14 ObjectType-Article-1 ObjectType-Feature-2 content type line 23
ISSN:	0277-6715 1097-0258 1097-0258
DOI:	10.1002/sim.5946