Fetching the paper…
Reading the bibliography…
We seek to learn models that we can interact with using high-level concepts: if the model did not think there was a bone spur in the x-ray, would it still predict severe arthritis? State-of-the-art models today do not typically support the manipulation of concepts like "the existence of bone spurs", as they are trained end-to-end to go directly from raw input (e.g., pixels) to output (e.g., arthritis severity).
Radiological assessment of osteo-arthrosis
Kellgren, J. and Lawrence, J · 1957
Earlier work this paper cites.
Feature selection and feature extract ion for text categorization
Lewis, D. D · 1992
Earlier work this paper cites.
Causality: Models, Reasoning and Inference , volume 29
Pearl, J · 2000
Earlier work this paper cites.
Kernel methods for relation extraction
Zelenko, D., Aone, C., and Richardella, A · 2003
Earlier work this paper cites.
A shortest path dependency kernel for relation extraction
Bunescu, R. C. and Mooney, R. J · 2005
Earlier work this paper cites.
Joint parsing and semantic role labeling
Sutton, C. and McCallum, A · 2005
Earlier work this paper cites.
The Osteoarthritis Initiative
Nevitt, M., Felson, D. T., and Lester, G · 2006
Earlier work this paper cites.
Attribute and simile classifiers for face verification
Kumar, N., Berg, A. C., Belhumeur, P. N., and Nayar, S. K · 2009
Earlier work this paper cites.
Learning to detect unseen object classes by between-class attribute transfer
Lampert, C. H., Nickisch, H., and Harmeling, S · 2009
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 dataset
Wah, C., Branson, S., Welinder, P., Perona, P., and Belongie, S · 2011
Earlier work this paper cites.
Discovering localized attributes for fine-grained recognition
Duan, K., Parikh, D., Crandall, D., and Grauman, K · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Bengio, Y., Courville, A., and Vincent, P · 2013
Earlier work this paper cites.
Flock: Hybrid Crowd-Machine learning classifiers
Cheng, J. and Bernstein, M. S · 2015
Earlier work this paper cites.
InfoGAN: Interpretable representation learning by information maximizing generative adversarial nets
Chen, X., Duan, Y., Houthooft, R., Schulman, J., Sutskever, I., and Abbeel, P · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Cited alongside, same era.
Part-stacked CNN for fine-grained visual categorization
Huang, S., Xu, Z., Tao, D., and Zhang, Y · 2016
Cited alongside, same era.
Rethinking the Inception architecture for computer vision
Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., and Wojna, Z · 2016
Cited alongside, same era.
Network dissection: Quantifying interpretability of deep visual representations
Bau, D., Zhou, B., Khosla, A., Oliva, A., and Torralba, A · 2017
Cited alongside, same era.
Interpretable explanations of black boxes by meaningful perturbation
Fong, R. C. and Vedaldi, A · 2017
Cited alongside, same era.
beta-vae: Learning basic visual concepts with a constrained variational framework
Higgins, I., Matthey, L., Pal, A., Burgess, C., Glorot, X., Botvinick, M., Mohamed, S., and Lerchner, A · 2017
Neural-symbolic vqa: Disentangling reasoning from vision and language understanding
Yi, K., Wu, J., Gan, C., Torralba, A., Kohli, P., and Tenenbaum, J · 2018
Later among the works it cites.
Feature engineering for machine learning: principles and techniques for data scientists
Zheng, A. and Casari, A · 2018
Later among the works it cites.
Interpretable basis decomposition for visual explanation
Zhou, B., Sun, Y., Bau, D., and Torralba, A · 2018
Later among the works it cites.
Global and local interpretability for cardiac MRI classification
Clough, J. R., Oksuz, I., Puyol-Antón, E., Ruijsink, B., King, A. P., and Schnabel, J. A · 2019
Later among the works it cites.
Towards automatic concept-based explanations
Ghorbani, A., Wexler, J., Zou, J. Y., and Kim, B · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Places: A 10 million image database for scene recognition
Zhou, B., Lapedriza, A., Khosla, A., Oliva, A., and Torralba, A · 2017
Cited alongside, same era.
Semantic bottleneck for computer vision tasks
Bucher, M., Herbin, S., and Jurie, F · 2018
Cited alongside, same era.
Large scale fine-grained categorization and domain-specific transfer learning
Cui, Y., Song, Y., Sun, C., Howard, A., and Belongie, S · 2018
Cited alongside, same era.
Clinically applicable deep learning for diagnosis and referral in retinal disease
Fauw, J. D., Ledsam, J. R., Romera-Paredes, B., Nikolov, S., Tomasev, N., Blackwell, S., Askham, H., Glorot, X., O’Donoghue, B., Visentin, D., et al · 2018
Cited alongside, same era.
Regression concept vectors for bidirectional explanations in histopathology
Graziani, M., Andrearczyk, V., and Müller, H · 2018
Cited alongside, same era.
Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)
Kim, B., Wattenberg, M., Gilmer, J., Cai, C., Wexler, J., Viegas, F., et al · 2018
Cited alongside, same era.
Goyal, Y., Shalit, U., and Kim, B · 2019
Later among the works it cites.
Interpretability beyond classification output: Semantic bottleneck networks
Losch, M., Fritz, M., and Schiele, B · 2019
Later among the works it cites.
Feature extraction and image processing for computer vision
Nixon, M. and Aguado, A · 2019
Later among the works it cites.
Using machine learning to understand racial and socioeconomic differences in knee pain
Pierson, E., Cutler, D., Leskovec, J., Mullainathan, S., and Obermeyer, Z · 2019
Later among the works it cites.
Interpretable AI for deep learning-based meteorological applications
Sprague, C., Wendoloski, E. B., and Guch, I · 2019
Later among the works it cites.
Concept whitening for interpretable image recognition
Chen, Z., Bei, Y., and Rudin, C · 2020
Closest in time.
ExpBERT: Representation engineering with natural language explanations
Murty, S., Koh, P. W., and Liang, P · 2020
Closest in time.
Generative causal explanations of black-box classifiers
O’Shaughnessy, M., Canal, G., Connor, M., Davenport, M., and Rozell, C · 2020
Closest in time.
Distributionally robust neural networks for group shifts: On the importance of regularization for worst-case generalization
Sagawa, S., Koh, P. W., Hashimoto, T. B., and Liang, P · 2020
Closest in time.