Fetching the paper…
Reading the bibliography…
The issue of disagreements amongst human experts is a ubiquitous one in both machine learning and medicine.
The proof and measurement of association between two things
Spearman, C · 1904
Earlier work this paper cites.
Maximum likelihood estimation of observer error-rates using the em algorithm
Dawid, P., Skene, A. M., Dawidt, A. P., and Skene, A. M · 1979
Earlier work this paper cites.
Regression modelling strategies for improved prognostic prediction
Harrell Jr, F. E., Lee, K. L., Califf, R. M., Pryor, D. B., and Rosati, R. A · 1984
Earlier work this paper cites.
Agreement among optometrists, ophthalmologists, and residents in evaluating the optic disc for glaucoma
Abrams, L. S., Scott, I. U., Spaeth, G. L., Quigley, H. A., and Varma, R · 1994
Earlier work this paper cites.
International Clinical Diabetic Retinopathy Disease Severity Scale Detailed Table
AAO · 2002
Earlier work this paper cites.
Toman’s tuberculosis. Case detection, treatment and monitoring: questions and answers
Daniel, T. M · 2004
Earlier work this paper cites.
Get another label? improving data quality and data mining using multiple, noisy labelers
Sheng, V. S., Provost, F., and Ipeirotis, P. G · 2008
Earlier work this paper cites.
Attribute learning in large-scale datasets
Russakovsky, O. and Fei-Fei, L · 2010
Earlier work this paper cites.
Online crowdsourcing: Rating annotators and obtaining cost-effective labels
Welinder, P. and Perona, P · 2010
Earlier work this paper cites.
Bayesian bias mitigation for crowdsourcing
Wauthier, F. L. and Jordan, M. I · 2011
Earlier work this paper cites.
Learning to label aerial images from noisy data
Mnih, V. and Hinton, G · 2012
Earlier work this paper cites.
Learning with noisy labels
Natarajan, N., Dhillon, I. S., Ravikumar, P. K., and Tewari, A · 2013
Earlier work this paper cites.
Training deep neural networks on noisy labels with bootstrapping
Reed, S. E., Lee, H., Anguelov, D., Szegedy, C., Erhan, D., and Rabinovich, A · 2014
Cited alongside, same era.
Diabetic retinopathy – biomolecules and multiple pathophysiology
Ahsan, H · 2015
Cited alongside, same era.
Diagnostic concordance among pathologists interpreting breast biopsy specimens
Elmore, J. G., Longton, G. M., Carney, P. A., Geller, B. M., Onega, T., Tosteson, A. N., Nelson, H. D., Pepe, M. S., Allison, K. H., Schnitt, S. J., et al · 2015
Cited alongside, same era.
On-the-job learning with bayesian decision theory
Werling, K., Chaganty, A. T., Liang, P. S., and Manning, C. D · 2015
Cited alongside, same era.
Learning from massive noisy labeled data for image classification
Xiao, T., Xia, T., Yang, Y., Huang, C., and Wang, X · 2015
Cited alongside, same era.
Axiomatic attribution for deep networks
Sundararajan, M., Taly, A., and Yan, Q · 2017
Later among the works it cites.
Bayesian image quality transfer with cnns: Exploring uncertainty in dmri super-resolution
Tanno, R., Worrall, D., Ghosh, A., Kaden, E., N. Sotiropoulos, S., Criminisi, A., and C. Alexander, D · 2017
Later among the works it cites.
Extent of diagnostic agreement among medical referrals
Van Such, M., Lohr, R., Beckman, T., and Naessens, J. M · 2017
Later among the works it cites.
Learning from noisy large-scale datasets with minimal supervision
Veit, A., Alldrin, N., Chechik, G., Krasin, I., Gupta, A., and Belongie, S. J · 2017
Later among the works it cites.
Why is my classifier discriminatory?
Chen, I., Johansson, F. D., and Sontag, D · 2018
Closest in time.
Clinically applicable deep learning for diagnosis and referral in retinal disease
De Fauw, J., Ledsam, J. R., Romera-Paredes, B., Nikolov, S., Tomasev, N., Blackwell, S., Askham, H., Glorot, X., O’Donoghue, B., Visentin, D., et al · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Development and validation of a deep learning algorithm for detection of diabetic retinopathy in retinal fundus photographs
Gulshan, V., Peng, L., Coram, M., Stumpe, M. C., Wu, D., Narayanaswamy, A., Venugopalan, S., Widner, K., Madams, T., Cuadros, J., Kim, R., Raman, R., Nelson, P. Q., Mega, J., and Webster, D · 2016
Cited alongside, same era.
On calibration of modern neural networks
Guo, C., Pleiss, G., Sun, Y., and Weinberger, K. Q · 2017
Cited alongside, same era.
What uncertainties do we need in bayesian deep learning for computer vision?
Kendall, A. and Gal, Y · 2017
Cited alongside, same era.
Chexnet: Radiologist-level pneumonia detection on chest x-rays with deep learning
Rajpurkar, P., Irvin, J., Zhu, K., Yang, B., Mehta, H., Duan, T., Ding, D., Bagul, A., Langlotz, C., Shpanskaya, K., Lungren, M. P., and Ng, A. Y · 2017
Cited alongside, same era.
Deep learning is robust to massive label noise
Rolnick, D., Veit, A., Belongie, S. J., and Shavit, N · 2017
Cited alongside, same era.
Smoothgrad: removing noise by adding noise
Smilkov, D., Thorat, N., Kim, B., Viégas, F., and Wattenberg, M · 2017
Cited alongside, same era.
Closest in time.
Who said what: Modeling individual labelers improves classification, 2018
Guan, M. Y., Gulshan, V., Dai, A. M., and Hinton, G. E · 2018
Closest in time.
Predicting foreground object ambiguity and efficiently crowdsourcing the segmentation (s)
Gurari, D., He, K., Xiong, B., Zhang, J., Sameki, M., Jain, S. D., Sclaroff, S., Betke, M., and Grauman, K · 2018
Closest in time.
A probabilistic u-net for segmentation of ambiguous images
Kohl, S. A., Romera-Paredes, B., Meyer, C., De Fauw, J., Ledsam, J. R., Maier-Hein, K. H., Eslami, S., Rezende, D. J., and Ronneberger, O · 2018
Closest in time.
Grader variability and the importance of reference standards for evaluating machine learning models for diabetic retinopathy
Krause, J., Gulshan, V., Rahimy, E., Karth, P., Widner, K., Corrado, G. S., Peng, L., and Webster, D. R · 2018
Closest in time.
Training convolutional networks with noisy labels
Sukhbaatar, S., Bruna, J., Paluri, M., Bourdev, L., and Fergus, R · 2080
Closest in time.