Fetching the paper…
Reading the bibliography…
As advancements in novel biomarker-based algorithms and models accelerate disease risk prediction and stratification in medicine, it is crucial to evaluate these models within the context of their intended clinical application.
Verification of forecasts expressed in terms of probability
Brier, G. W. (1950) · 1950
Earlier work this paper cites.
A new vector partition of the probability score
Murphy, A. H. (1973) · 1973
Earlier work this paper cites.
Therapeutic decision making: a cost-benefit analysis
Pauker, S. G. and Kassirer, J. P. (1975) · 1975
Earlier work this paper cites.
Goodness of fit tests for the multiple logistic regression model
Hosmer, D. W. and Lemesbow, S. (1980) · 1980
Earlier work this paper cites.
Probabilistic prediction in patient management and clinical trials
Spiegelhalter, D. J. (1986) · 1986
Earlier work this paper cites.
A general method for comparing probability assessors
Schervish, M. J. (1989) · 1989
Earlier work this paper cites.
Bias and efficiency loss due to misclassified responses in binary regression
Neuhaus, J. M. (1999) · 1999
Earlier work this paper cites.
The foundations of cost-sensitive learning
Elkan, C. (2001) · 2001
Earlier work this paper cites.
Combining several screening tests: optimality of the risk score
McIntosh, M. W. and Pepe, M. S. (2002) · 2002
Earlier work this paper cites.
The statistical evaluation of medical tests for classification and prediction
Pepe, M. S. (2003) · 2003
Earlier work this paper cites.
On criteria for evaluating models of absolute risk
Gail, M. H. and Pfeiffer, R. M. (2005) · 2005
Earlier work this paper cites.
Survival model predictive accuracy and roc curves
Heagerty, P. J. and Zheng, Y. (2005) · 2005
Earlier work this paper cites.
Decision curve analysis: a novel method for evaluating prediction models
Vickers, A. J. and Elkin, E. B. (2006) · 2006
Earlier work this paper cites.
Strictly proper scoring rules, prediction, and estimation
Gneiting, T. and Raftery, A. E. (2007) · 2007
Earlier work this paper cites.
The performance of risk prediction models
Gerds, T. A., Cai, T., and Schumacher, M. (2008) · 2008
Earlier work this paper cites.
Reliability, sufficiency, and the decomposition of proper scores
Bröcker, J. (2009) · 2009
Earlier work this paper cites.
Measuring classifier performance: a coherent alternative to the area under the roc curve
Hand, D. J. (2009) · 2009
Earlier work this paper cites.
Evaluating diagnostic tests: the area under the roc curve and the balance of errors
Hand, D. J. (2010) · 2010
Earlier work this paper cites.
Weighted area under the receiver operating characteristic curve and its application to gene selection
Li, J. and Fine, J. P. (2010) · 2010
Cited alongside, same era.
Composite binary losses
Reid, M. D. and Williamson, R. C. (2010) · 2010
Cited alongside, same era.
Use of brier score to assess binary predictions
Rufibach, K. (2010) · 2010
Cited alongside, same era.
Assessing the performance of prediction models: a framework for some traditional and novel measures
Steyerberg, E. W., Vickers, A. J., Cook, N. R., Gerds, T., Gonen, M., Obuchowski, N., Pencina, M. J., and Kattan, M. W. (2010) · 2010
Cited alongside, same era.
A regret theory approach to decision curve analysis: a novel method for eliciting decision makers’ preferences and decision-making
Tsalatsanis, A., Hozo, I., Vickers, A., and Djulbegovic, B. (2010) · 2010
Cited alongside, same era.
Information, divergence and risk for binary experiments
Beyond discrimination: a comparison of calibration methods and clinical usefulness of predictive models of readmission risk
Walsh, C. G., Sharman, K., and Hripcsak, G. (2017) · 2017
Later among the works it cites.
The index of prediction accuracy: an intuitive measure useful for evaluating risk prediction models
Kattan, M. W. and Gerds, T. A. (2018) · 2018
Later among the works it cites.
The c-index is not proper for the evaluation of-year predicted risks
Blanche, P., Kattan, M. W., and Gerds, T. A. (2019) · 2019
Later among the works it cites.
Score decompositions in forecast verification
Mitchell, K. (2019) · 2019
Later among the works it cites.
Calibration: the achilles heel of predictive analytics
Van Calster, B., McLernon, D. J., Van Smeden, M., Wynants, L., and Steyerberg, E. W. (2019) · 2019
Later among the works it cites.
Measuring the temporal prognostic utility of a baseline risk score
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Reid, M. and Williamson, R. (2011) · 2011
Cited alongside, same era.
Probabilistic forecasting
Gneiting, T. and Katzfuss, M. (2014) · 2014
Cited alongside, same era.
A note on the evaluation of novel biomarkers: do not rely on integrated discrimination improvement and net reclassification index
Hilden, J. and Gerds, T. A. (2014) · 2014
Cited alongside, same era.
Regression modeling strategies: with applications to linear models, logistic and ordinal regression, and survival analysis
Harrell Jr, F. E. (2015) · 2015
Cited alongside, same era.
Benchmarking state-of-the-art classification algorithms for credit scoring: An update of research
Lessmann, S., Baesens, B., Seow, H.-V., and Thomas, L. C. (2015) · 2015
Cited alongside, same era.
The net reclassification index (nri): a misleading measure of prediction improvement even with independent test data sets
Pepe, M. S., Fan, J., Feng, Z., Gerds, T., and Hilden, J. (2015) · 2015
Cited alongside, same era.
Stata base reference manual release 14
Stata, A., Publication, P., and Lp, S. (2015) · 2015
Cited alongside, same era.
Devlin, S. M., Gönen, M., and Heller, G. (2020) · 2020
Later among the works it cites.
Effective ways to build and evaluate individual survival distributions
Haider, H., Hoehn, B., Davis, S., and Greiner, R. (2020) · 2020
Later among the works it cites.
A tutorial on calibration measurements and calibration models for clinical prediction models
Huang, Y., Li, W., Macheret, F., Gabriel, R. A., and Ohno-Machado, L. (2020) · 2020
Later among the works it cites.
Statistical inference for net benefit measures in biomarker validation studies
Marsh, T. L., Janes, H., and Pepe, M. S. (2020) · 2020
Later among the works it cites.
Estimating the decision curve and its precision from three study designs
Pfeiffer, R. M. and Gail, M. H. (2020) · 2020
Later among the works it cites.
Statistical inference for decision curve analysis, with applications to cataract diagnosis
Sande, S. Z., Li, J., D’Agostino, R., Yin Wong, T., and Cheng, C.-Y. (2020) · 2020
Later among the works it cites.
Concordance probability as a meaningful contrast across disparate survival times
Devlin, S. M. and Heller, G. (2021) · 2021
Later among the works it cites.
Stable reliability diagrams for probabilistic classifiers
Dimitriadis, T., Gneiting, T., and Jordan, A. I. (2021) · 2021
Later among the works it cites.
A review on fairness in machine learning
Pessach, D. and Shmueli, E. (2022) · 2022
Later among the works it cites.
Survival regression with proper scoring rules and monotonic neural networks
Rindt, D., Hu, R., Steinsaltz, D., and Sejdinovic, D. (2022) · 2022
Later among the works it cites.
Model diagnostics and forecast evaluation for quantiles
Gneiting, T., Wolffram, D., Resin, J., Kraus, K., Bracher, J., Dimitriadis, T., Hagenmeyer, V., Jordan, A. I., Lerch, S., Phipps, K., et al. (2023) · 2023
Later among the works it cites.
An effective meaningful way to evaluate survival models
Qi, S.-a., Kumar, N., Farrokh, M., Sun, W., Kuan, L.-H., Ranganath, R., Henao, R., and Greiner, R. (2023) · 2023
Later among the works it cites.