Fetching the paper…
Reading the bibliography…
Data-centric AI calls for better, not just bigger, datasets.
Science and Statistics
G. Box · 1976
Earlier work this paper cites.
Neural Networks and the Bias/Variance Dilemma
S. Geman, E. Bienenstock, and R. Doursat · 1992
Earlier work this paper cites.
Bias in Computer Systems
B. Friedman and H. Nissenbaum · 1996
Earlier work this paper cites.
A Taxonomy of Privacy
D. J. Solove · 2006
Earlier work this paper cites.
Towards Scalable Dataset Construction: An Active Learning Approach
B. Collins, J. Deng, K. Li, and L. Fei-Fei · 2008
Earlier work this paper cites.
A Comprehensive Analysis of Information Leakage in Deep Transfer Learning
C. Chen, B. Wu, M. Qiu, L. Wang, and J. Zhou · 2009
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Turkopticon: interrupting worker invisibility in amazon mechanical turk
L. C. Irani and M. S. Silberman · 2013
Earlier work this paper cites.
Hidden Technical Debt in Machine Learning Systems
D. Sculley, G. Holt, D. Golovin, E. Davydov, T. Phillips, D. Ebner, V. Chaudhary, M. Young, J.-F. Crespo, and D. Dennison · 2015
Earlier work this paper cites.
Ten simple rules for responsible big data research
M. Zook, S. Barocas, D. Boyd, K. Crawford, E. Keller, S. P. Gangadharan, A. Goodman, R. Hollander, B. A. Koenig, J. Metcalf, A. Narayanan, A. Nelson, and F. Pasquale · 2017
Earlier work this paper cites.
Delimiting the concept of personal data after the GDPR
B. Wong · 2018
Earlier work this paper cites.
A Kernel Theory of Modern Data Augmentation
T. Dao, A. Gu, A. J. Ratner, V. Smith, C. D. Sa, and C. Re · 2019
Cited alongside, same era.
Big Data and Discrimination
T. B. Gillis and J. L. Spiess · 2019
Cited alongside, same era.
The Research Exemption Carve Out: Understanding Research Participants Rights Under Gdpr and U.s. Data Privacy Laws
C. Mabel and S. Tara · 2019
Cited alongside, same era.
Bias In, Bias Out
S. G. Mayson · 2019
Cited alongside, same era.
Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks against Centralized and Federated Learning
M. Nasr, R. Shokri, and A. Houmansadr · 2019
Cited alongside, same era.
The New Legal Landscape for Text Mining and Machine Learning
M. Sag · 2019
Cited alongside, same era.
Data and its (dis)contents: A survey of dataset development and use in machine learning research
A. Paullada, I. D. Raji, E. M. Bender, E. Denton, and A. Hanna · 2020
Later among the works it cites.
The relationship between trust in ai and trustworthy machine learning technologies
E. Toreini, M. Aitken, K. Coopamootoo, K. Elliott, C. G. Zelaya, and A. van Moorsel · 2020
Later among the works it cites.
Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure
B. Hutchinson, A. Smart, A. Hanna, E. Denton, C. Greer, O. Kjartansson, P. Barnes, and M. Mitchell · 2021
Closest in time.
Researchers Blur Faces That Launched a Thousand Algorithms
W. Knight · 2021
Closest in time.
Facebook to shut down face-recognition system, delete data
M. O’Brien and B. Ortutay · 2021
Closest in time.
Just What do You Think You’re Doing, Dave?’ A Checklist for Responsible Data Use in NLP
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Data protection in the age of big data
S. Wachter · 2019
Cited alongside, same era.
Attachment and trust in artificial intelligence
O. Gillath, T. Ai, M. S. Branicky, S. Keshmiri, R. B. Davison, and R. Spaulding · 2020
Cited alongside, same era.
Measuring Algorithmic Fairness
D. Hellman · 2020
Cited alongside, same era.
Pretrained Transformers Improve Out-of-Distribution Robustness
D. Hendrycks, X. Liu, E. Wallace, A. Dziedzic, R. Krishnan, and D. Song · 2020
Cited alongside, same era.
Before and beyond trust: reliance in medical AI
C. X. Kerasidou, A. Kerasidou, M. Buscher, and S. Wilkinson · 2020
Cited alongside, same era.
A. Rogers, T. Baldwin, and K. Leins · 2021
Closest in time.
The right to privacy in the digital age
United Nations High Commissioner for Human Rights · 2021
Closest in time.
Implications of Data Anonymization on the Statistical Evidence of Disparity
H. Xu and N. Zhang · 2021
Closest in time.
Are Larger Pretrained Language Models Uniformly Better? Comparing Performance at the Instance Level
R. Zhong, D. Ghosh, D. Klein, and J. Steinhardt · 2021
Closest in time.
Estimating the success of re-identifications in incomplete datasets using generative models
L. Rocher, J. M. Hendrickx, and Y.-A. de Montjoye · 2041
Closest in time.
The trouble with European data protection law
B.-J. Koops · 2044
Closest in time.