Confidence intervals of prediction accuracy measures for multivariable prediction models based on the bootstrap-based optimism correction methods

Noma, Hisashi; Shinozaki, Tomohiro; Iba, Katsuhiro; Teramukai, Satoshi; Furukawa, Toshi A.

Statistics > Methodology

arXiv:2005.01457v4 (stat)

[Submitted on 4 May 2020 (v1), revised 30 Jun 2021 (this version, v4), latest version 25 Jul 2021 (v5)]

Title:Confidence intervals of prediction accuracy measures for multivariable prediction models based on the bootstrap-based optimism correction methods

Authors:Hisashi Noma, Tomohiro Shinozaki, Katsuhiro Iba, Satoshi Teramukai, Toshi A. Furukawa

View PDF

Abstract:In assessing prediction accuracy of multivariable prediction models, optimism corrections are essential for preventing biased results. However, in most published papers of clinical prediction models, the point estimates of the prediction accuracy measures are corrected by adequate bootstrap-based correction methods, but their confidence intervals are not corrected, e.g., the DeLong's confidence interval is usually used for assessing the C-statistic. These naive methods do not adjust for the optimism bias and do not account for statistical variability in the estimation of parameters in the prediction models. Therefore, their coverage probabilities of the true value of the prediction accuracy measure can be seriously below the nominal level (e.g., 95%). In this article, we provide two generic bootstrap methods, namely (1) location-shifted bootstrap confidence intervals and (2) two-stage bootstrap confidence intervals, that can be generally applied to the bootstrap-based optimism correction methods, i.e., the Harrell's bias correction, 0.632, and 0.632+ methods. In addition, they can be widely applied to various methods for prediction model development involving modern shrinkage methods such as the ridge and lasso regressions. Through numerical evaluations by simulations, the proposed confidence intervals showed favourable coverage performances. Besides, the current standard practices based on the optimism-uncorrected methods showed serious undercoverage properties. To avoid erroneous results, the optimism-uncorrected confidence intervals should not be used in practice, and the adjusted methods are recommended instead. We also developed the R package predboot for implementing these methods (this https URL). The effectiveness of the proposed methods are illustrated via applications to the GUSTO-I clinical trial.

Subjects:	Methodology (stat.ME); Applications (stat.AP); Computation (stat.CO)
Cite as:	arXiv:2005.01457 [stat.ME]
	(or arXiv:2005.01457v4 [stat.ME] for this version)
	https://doi.org/10.48550/arXiv.2005.01457

Submission history

From: Hisashi Noma [view email]
[v1] Mon, 4 May 2020 13:15:54 UTC (290 KB)
[v2] Tue, 5 May 2020 11:33:59 UTC (289 KB)
[v3] Fri, 8 May 2020 08:40:44 UTC (289 KB)
[v4] Wed, 30 Jun 2021 10:18:57 UTC (395 KB)
[v5] Sun, 25 Jul 2021 12:56:27 UTC (393 KB)

Statistics > Methodology

Title:Confidence intervals of prediction accuracy measures for multivariable prediction models based on the bootstrap-based optimism correction methods

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Methodology

Title:Confidence intervals of prediction accuracy measures for multivariable prediction models based on the bootstrap-based optimism correction methods

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators