Fetching the paper…
Reading the bibliography…
Machine learning models are often brittle on production data despite achieving high accuracy on benchmark datasets.
1907
Earlier work this paper cites.
1910
Earlier work this paper cites.
S. Kerr, “On the folly of rewarding a, while hoping for b,” Academy of Management journal , vol. 18, no. 4, pp. 769–783, 1975
1975
Earlier work this paper cites.
D. E. Forsythe, “Engineering knowledge: The construction of knowledge in artificial intelligence,” Social studies of science , vol. 23, no. 3, pp. 445–477, 1993
1993
Earlier work this paper cites.
W. Ickes, “Empathic accuracy,” Journal of personality , vol. 61, no. 4, pp. 587–610, 1993
1993
Earlier work this paper cites.
G. Widmer and M. Kubat, “Learning in the presence of concept drift and hidden contexts,” Machine learning , vol. 23, no. 1, pp. 69–101, 1996
1996
Earlier work this paper cites.
R. W. Picard, Affective computing , 2000
2000
Earlier work this paper cites.
2003
Earlier work this paper cites.
2003
Earlier work this paper cites.
M. Pantic, M. Valstar, R. Rademaker, and L. Maat, “Web-based database for facial expression analysis,” in 2005 IEEE international conference on multimedia and Expo . IEEE, 2005, pp. 5–pp
2005
Earlier work this paper cites.
2006
Earlier work this paper cites.
J. Quiñonero-Candela, M. Sugiyama, A. Schwaighofer, and N. D. Lawrence, Dataset shift in machine learning , 2008
2008
Earlier work this paper cites.
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in 2009 IEEE conference on computer vision and pattern recognition . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
2009
Earlier work this paper cites.
T. Bänziger and K. R. Scherer, “Introducing the geneva multimodal emotion portrayal (gemep) corpus,” Blueprint for affective computing: A sourcebook , vol. 2010, pp. 271–94, 2010
2010
Earlier work this paper cites.
P. Lucey, J. F. Cohn, T. Kanade, J. Saragih, Z. Ambadar, and I. Matthews, “The extended cohn-kanade dataset (ck+): A complete dataset for action unit and emotion-specified expression,” in 2010 ieee computer society conference on computer vision and pattern recognition-workshops . IEEE, 2010, pp. 94–101
2010
Earlier work this paper cites.
R. Gross, I. Matthews, J. Cohn, T. Kanade, and S. Baker, “Multi-pie,” Image and vision computing , vol. 28, no. 5, pp. 807–813, 2010
2010
Earlier work this paper cites.
A. Torralba and A. A. Efros, “Unbiased look at dataset bias,” in CVPR 2011 . IEEE, 2011, pp. 1521–1528
2011
Earlier work this paper cites.
A. Dhall, R. Goecke, S. Lucey, and T. Gedeon, “Static facial expression analysis in tough conditions: Data, evaluation protocol and benchmark,” in 2011 IEEE International Conference on Computer Vision Workshops (ICCV Workshops) . IEEE, 2011, pp. 2106–2112
2011
Earlier work this paper cites.
M. F. Valstar, B. Jiang, M. Mehu, M. Pantic, and K. Scherer, “The first facial expression recognition and analysis challenge,” in 2011 IEEE International Conference on Automatic Face & Gesture Recognition (FG) . IEEE, 2011, pp. 921–926
2011
Earlier work this paper cites.
L. Deng, “The mnist database of handwritten digit images for machine learning research [best of the web],” IEEE signal processing magazine , vol. 29, no. 6, pp. 141–142, 2012
2012
Earlier work this paper cites.
J. G. Moreno-Torres, T. Raeder, R. Alaiz-Rodríguez, N. V. Chawla, and F. Herrera, “A unifying view on dataset shift in classification,” Pattern recognition , vol. 45, no. 1, pp. 521–530, 2012
2012
Earlier work this paper cites.
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” Advances in neural information processing systems , vol. 25, 2012
2012
Earlier work this paper cites.
S. M. Mavadati, M. H. Mahoor, K. Bartlett, P. Trinh, and J. F. Cohn, “Disfa: A spontaneous facial action intensity database,” IEEE Transactions on Affective Computing , vol. 4, no. 2, pp. 151–160, 2013
2013
Earlier work this paper cites.
I. J. Goodfellow, D. Erhan, P. L. Carrier, A. Courville, M. Mirza, B. Hamner, W. Cukierski, Y. Tang, D. Thaler, D.-H. Lee et al. , “Challenges in representation learning: A report on three machine learning contests,” in International conference on neural information processing . Springer, 2013, pp. 117–124
2013
Earlier work this paper cites.
A. Krizhevsky, V. Nair, and G. Hinton, “The cifar-10 dataset,” online: http://www. cs. toronto. edu/kriz/cifar. html , vol. 55, no. 5, 2014
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “Imagenet large scale visual recognition challenge,” International journal of computer vision , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
J. F. Cohn and F. De la Torre, “Automated face analysis for affective computing.” 2015
2015
Earlier work this paper cites.
A. E. Johnson, T. J. Pollard, L. Shen, L.-w. H. Lehman, M. Feng, M. Ghassemi, B. Moody, P. Szolovits, L. Anthony Celi, and R. G. Mark, “Mimic-iii, a freely accessible critical care database,” Scientific data , vol. 3, no. 1, pp. 1–9, 2016
2016
Earlier work this paper cites.
A. Mollahosseini, D. Chan, and M. H. Mahoor, “Going deeper in facial expression recognition using deep neural networks,” in 2016 IEEE Winter conference on applications of computer vision (WACV) . IEEE, 2016, pp. 1–10
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Mullainathan and Z. Obermeyer, “Does Machine Learning Automate Moral Hazard and Error?” vol. 107, no. 5, p. 5, 2017
2017
Cited alongside, same era.
A. Paiva, I. Leite, H. Boukricha, and I. Wachsmuth, “Empathy in virtual agents and robots: A survey,” ACM Transactions on Interactive Intelligent Systems (TiiS) , vol. 7, no. 3, pp. 1–40, 2017
2017
Cited alongside, same era.
A. Esteva, B. Kuprel, R. A. Novoa, J. Ko, S. M. Swetter, H. M. Blau, and S. Thrun, “Dermatologist-level classification of skin cancer with deep neural networks,” Nature , vol. 542, no. 7639, pp. 115–118, Feb. 2017. [Online]. Available: http://www.nature.com/articles/nature21056
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2021
Later among the works it cites.
M. Groh, C. Harris, L. Soenksen, F. Lau, R. Han, A. Kim, A. Koochek, and O. Badri, “Evaluating Deep Neural Networks Trained on Clinical Images in Dermatology with the Fitzpatrick 17k Dataset,” in 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) . Nashville, TN, USA: IEEE, Jun. 2021, pp. 1820–1828. [Online]. Available: https://ieeexplore.ieee.org/document/9522867/
2021
Later among the works it cites.
B. Schölkopf, F. Locatello, S. Bauer, N. R. Ke, N. Kalchbrenner, A. Goyal, and Y. Bengio, “Toward causal representation learning,” Proceedings of the IEEE , vol. 109, no. 5, pp. 612–634, 2021
2021
Later among the works it cites.
T. Gebru, J. Morgenstern, B. Vecchione, J. W. Vaughan, H. Wallach, H. D. Iii, and K. Crawford, “Datasheets for datasets,” Communications of the ACM , vol. 64, no. 12, pp. 86–92, 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. C. Lipton, “The mythos of model interpretability,” Communications of the ACM , vol. 61, no. 10, pp. 36–43, Sep. 2018. [Online]. Available: https://dl.acm.org/doi/10.1145/3233231
2018
Cited alongside, same era.
J. Buolamwini and T. Gebru, “Gender shades: Intersectional accuracy disparities in commercial gender classification,” in Conference on fairness, accountability and transparency . PMLR, 2018, pp. 77–91
2018
Cited alongside, same era.
K. Eykholt, I. Evtimov, E. Fernandes, B. Li, A. Rahmati, C. Xiao, A. Prakash, T. Kohno, and D. Song, “Robust physical-world attacks on deep learning visual classification,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 1625–1634
2018
Cited alongside, same era.
2018
Cited alongside, same era.
E. M. Bender and B. Friedman, “Data statements for natural language processing: Toward mitigating system bias and enabling better science,” Transactions of the Association for Computational Linguistics , vol. 6, pp. 587–604, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
J. Irvin, P. Rajpurkar, M. Ko, Y. Yu, S. Ciurea-Ilcus, C. Chute, H. Marklund, B. Haghgoo, R. Ball, K. Shpanskaya et al. , “Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,” in Proceedings of the AAAI conference on artificial intelligence , vol. 33, no. 01, 2019, pp. 590–597
2019
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
S. G. Finlayson, A. Subbaswamy, K. Singh, J. Bowers, A. Kupke, J. Zittrain, I. S. Kohane, and S. Saria, “The Clinician and Dataset Shift in Artificial Intelligence,” New England Journal of Medicine , vol. 385, no. 3, pp. 283–286, Jul. 2021. [Online]. Available: http://www.nejm.org/doi/10.1056/NEJMc2104626
2021
Later among the works it cites.
E. Pierson, D. M. Cutler, J. Leskovec, S. Mullainathan, and Z. Obermeyer, “An algorithmic approach to reducing unexplained pain disparities in underserved populations,” Nature Medicine , vol. 27, no. 1, pp. 136–140, Jan. 2021. [Online]. Available: http://www.nature.com/articles/s41591-020-01192-7
2021
Later among the works it cites.
D. Kerrigan, J. Hullman, and E. Bertini, “A survey of domain knowledge elicitation in applied machine learning,” Multimodal Technologies and Interaction , vol. 5, no. 12, p. 73, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Mullainathan and Z. Obermeyer, “On the Inequity of Predicting A While Hoping for B,” AEA Papers and Proceedings , vol. 111, pp. 37–42, May 2021. [Online]. Available: https://pubs.aeaweb.org/doi/10.1257/pandp.20211078
2021
Later among the works it cites.
A. Paullada, I. D. Raji, E. M. Bender, E. Denton, and A. Hanna, “Data and its (dis) contents: A survey of dataset development and use in machine learning research,” Patterns , vol. 2, no. 11, p. 100336, 2021
2021
Later among the works it cites.
A. S. Cowen, D. Keltner, F. Schroff, B. Jou, H. Adam, and G. Prasad, “Sixteen facial expressions occur in similar contexts worldwide,” Nature , vol. 589, no. 7841, pp. 251–257, Jan. 2021. [Online]. Available: http://www.nature.com/articles/s41586-020-3037-7
2021
Later among the works it cites.
A. Sankaranarayanan, M. Groh, R. Picard, and A. Lippman, “The presidential deepfakes dataset,” 2021
2021
Later among the works it cites.
A. Jain, D. Way, V. Gupta, Y. Gao, G. de Oliveira Marinho, J. Hartford, R. Sayres, K. Kanada, C. Eng, K. Nagpal, K. B. DeSalvo, G. S. Corrado, L. Peng, D. R. Webster, R. C. Dunn, D. Coz, S. J. Huang, Y. Liu, P. Bui, and Y. Liu, “Development and Assessment of an Artificial Intelligence–Based Tool for Skin Condition Diagnosis by Primary Care Physicians and Nurse Practitioners in Teledermatology Practices,” JAMA Network Open , vol. 4, no. 4, p. e217249, Apr. 2021. [Online]. Available: https://jamanetwork.com/journals/jamanetworkopen/fullarticle/2779250
2021
Later among the works it cites.
S. Gaube, H. Suresh, M. Raue, A. Merritt, S. J. Berkowitz, E. Lermer, J. F. Coughlin, J. V. Guttag, E. Colak, and M. Ghassemi, “Do as AI say: susceptibility in deployment of clinical decision-aids,” npj Digital Medicine , vol. 4, no. 1, p. 31, Dec. 2021. [Online]. Available: http://www.nature.com/articles/s41746-021-00385-9
2021
Later among the works it cites.
M. Jacobs, M. F. Pradier, T. H. McCoy, R. H. Perlis, F. Doshi-Velez, and K. Z. Gajos, “How machine-learning recommendations influence clinician treatment selections: the example of antidepressant selection,” Translational Psychiatry , vol. 11, no. 1, p. 108, Jun. 2021. [Online]. Available: http://www.nature.com/articles/s41398-021-01224-x
2021
Later among the works it cites.
R. Daneshjou, K. Vodrahalli, R. A. Novoa, M. Jenkins, W. Liang, V. Rotemberg, J. Ko, S. M. Swetter, E. E. Bailey, O. Gevaert et al. , “Disparities in dermatology ai performance on a diverse, curated clinical image set,” Science advances , vol. 8, no. 31, p. eabq6147, 2022
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
R. L. Thomas and D. Uminsky, “Reliance on metrics is a fundamental challenge for ai,” Patterns , vol. 3, no. 5, p. 100476, 2022
2022
Closest in time.
M. Makar, B. Packer, D. Moldovan, D. Blalock, Y. Halpern, and A. D’Amour, “Causally motivated shortcut removal using auxiliary labels,” in International Conference on Artificial Intelligence and Statistics . PMLR, 2022, pp. 739–766
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
M. C. Pelz, K. R. Allen, J. B. Tenenbaum, and L. E. Schulz, “Foundations of intuitive power analyses in children and adults,” Nature Human Behaviour , pp. 1–12, 2022
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
M. Groh, Z. Epstein, C. Firestone, and R. Picard, “Deepfake detection by human crowds, machines, and machine-informed crowds,” Proceedings of the National Academy of Sciences , vol. 119, no. 1, p. e2110013119, Jan. 2022. [Online]. Available: http://www.pnas.org/lookup/doi/10.1073/pnas.2110013119
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
A. Ilyas, S. M. Park, L. Engstrom, G. Leclerc, and A. Madry, “Datamodels: Understanding predictions with data and data with predictions,” in International Conference on Machine Learning . PMLR, 2022, pp. 9525–9587
2022
Closest in time.
J. J. Smith, S. Amershi, S. Barocas, H. Wallach, and J. Wortman Vaughan, “Real ml: Recognizing, exploring, and articulating limitations of machine learning research,” in 2022 ACM Conference on Fairness, Accountability, and Transparency , 2022, pp. 587–597
2022
Closest in time.
2022
Closest in time.
J. Wakefield, “Deepfake presidents used in russia-ukraine war,” Mar 2022. [Online]. Available: https://www.bbc.com/news/technology-60780142
2022
Closest in time.
2022
Closest in time.
Y. Assael, T. Sommerschield, B. Shillingford, M. Bordbar, J. Pavlopoulos, M. Chatzipanagiotou, I. Androutsopoulos, J. Prag, and N. de Freitas, “Restoring and attributing ancient texts using deep neural networks,” Nature , vol. 603, no. 7900, pp. 280–283, Mar. 2022. [Online]. Available: https://www.nature.com/articles/s41586-022-04448-z
2022
Closest in time.