Fetching the paper…
Reading the bibliography…
As machine learning models become increasingly larger, trained weakly supervised on large, possibly uncurated data sets, it becomes increasingly important to establish mechanisms for inspecting, interacting, and revising models to mitigate learning shortcuts and guarantee their learned knowledge is aligned with human knowledge.
Language Models are Few-Shot Learners
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, et al · 1901
Earlier work this paper cites.
Intelligence Through Interaction: Towards a Unified Theory for Learning
Tan AH, Carpenter GA, Grossberg S · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng J, Dong W, Socher R, Li LJ, Li K, Fei-Fei L · 2009
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma DP, Ba J · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Simonyan K, Zisserman A · 2015
Earlier work this paper cites.
Interpretation of Prediction Models Using the Input Gradient
Hechtlinger Y · 2016
Earlier work this paper cites.
”Why Should I Trust You?”: Explaining the Predictions of Any Classifier
Ribeiro MT, Singh S, Guestrin C · 2016
Earlier work this paper cites.
Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based Localization
Selvaraju RR, Das A, Vedantam R, Cogswell M, Parikh D, Batra D · 2017
Earlier work this paper cites.
Enlightening Deep Neural Networks with Knowledge of Confounding Factors
Zhong Y, Ettinger G · 2017
Earlier work this paper cites.
Right for the Right Reasons: Training Differentiable Models by Constraining their Explanations
Ross AS, Hughes MC, Doshi-Velez F · 2017
Earlier work this paper cites.
Towards A Rigorous Science of Interpretable Machine Learning
Doshi-Velez F, Kim B · 2017
Earlier work this paper cites.
Skin Lesion Analysis Toward Melanoma Detection: A Challenge at the 2017 International Symposium on Biomedical Imaging (ISBI), Hosted by the International Skin Imaging Collaboration (ISIC)
Codella N, Gutman D, Celebi ME, Helba B, Marchetti M, Dusza S, et al · 2017
Earlier work this paper cites.
Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
Xiao H, Rasul K, Vollgraf R · 2017
Earlier work this paper cites.
Sanity checks for saliency maps
Adebayo J, Gilmer J, Muelly M, Goodfellow I, Hardt M, Kim B · 2018
Earlier work this paper cites.
The HAM10000 dataset, a large collection of multi-source dermatoscopic images of common pigmented skin lesions
Tschandl P, Rosendahl C, Kittler H · 2018
Cited alongside, same era.
Dissecting racial bias in an algorithm used to manage the health of populations
Obermeyer Z, Powers B, Vogeli C, Mullainathan S · 2019
Cited alongside, same era.
Analysis Methods in Neural Language Processing: A Survey
Belinkov Y, Glass J · 2019
Cited alongside, same era.
Unmasking Clever Hans Predictors and Assessing What Machines Really Learn
Lapuschkin S, Wäldchen S, Binder A, Montavon G, Samek W, Müller K · 2019
Cited alongside, same era.
Explanatory Interactive Machine Learning
Teso S, Kersting K · 2019
Cited alongside, same era.
Taking a HINT: Leveraging Explanations to Make Vision and Language Models More Grounded
Selvaraju RR, Lee S, Shen Y, Jin H, Batra D, Parikh D · 2019
The Next Frontier: AI We Can Really Trust
Holzinger A · 2021
Later among the works it cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?
Bender EM, Gebru T, McMillan-Major A, Shmitchell S · 2021
Later among the works it cites.
Right for the Right Concept: Revising Neuro-Symbolic Concepts by Interacting with their Explanations
Stammer W, Schramowski P, Kersting K · 2021
Later among the works it cites.
Right for better reasons: Training differentiable models by constraining their influence function
Shao X, Skryagin A, Schramowski P, Stammer W, Kersting K · 2021
Later among the works it cites.
Cooperative AI: machines must learn to find common ground
Dafoe A, Bachrach Y, Hadfield G, Horvitz E, Larson K, Graepel T · 2021
Later among the works it cites.
Training a GAN To Explain a Classifier in StyleSpace
Lang O, Gandelsman Y, Yarom M, Wald Y, Elidan G, Hassidim A, et al · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Analyzing ImageNet with Spectral Relevance Analysis: Towards ImageNet un-Hans’ed
Anders CJ, Marinc T, Neumann D, Samek W, Müller K, Lapuschkin S · 2019
Cited alongside, same era.
BCN20000: Dermoscopic Lesions in the Wild
Combalia M, Codella NCF, Rotemberg V, Helba B, Vilaplana V, Reiter O, et al · 2019
Cited alongside, same era.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Paszke A, Gross S, Massa F, Lerer A, Bradbury J, Chanan G, et al · 2019
Cited alongside, same era.
Shortcut learning in deep neural networks
Geirhos R, Jacobsen JH, Michaelis C, Zemel R, Brendel W, Bethge M, et al · 2020
Cited alongside, same era.
A Diagnostic Study of Explainability Techniques for Text Classification
Atanasova P, Simonsen JG, Lioma C, Augenstein I · 2020
Cited alongside, same era.
Making deep neural networks right for the right scientific reasons by interacting with their explanations
Schramowski P, Stammer W, Teso S, Brugger A, Herbert F, Shao X, et al · 2020
Cited alongside, same era.
Later among the works it cites.
Oxford Dictionary; 2022
Trust; Definition and Meaning of trust on Lexico.com · 2022
Closest in time.
Hierarchical Text-Conditional Image Generation with CLIP Latents
Ramesh A, Dhariwal P, Nichol A, Chu C, Chen M · 2022
Closest in time.
Fairness and Explanation in AI-Informed Decision Making
Angerschmid A, Zhou J, Theuermann K, Chen F, Holzinger A · 2022
Closest in time.
Leveraging Explanations in Interactive Machine Learning: An Overview
Teso S, Alkan Ö, Stammer W, Daly E · 2022
Closest in time.
The Disagreement Problem in Explainable Machine Learning: A Practitioner’s Perspective
Krishna S, Han T, Gu A, Pombra J, Jabbari S, Wu S, et al · 2022
Closest in time.
CAIPI in Practice: Towards Explainable Interactive Medical Image Classification
Slany E, Ott Y, Scheele S, Paulus J, Schmid U · 2022
Closest in time.
A Typology to Explore the Mitigation of Shortcut Behavior
Friedrich F, Stammer W, Schramowski P, Kersting K · 2022
Closest in time.