Fetching the paper…
Reading the bibliography…
Structural analysis methods (e.g., probing and feature attribution) are increasingly important tools for neural network analysis.
Bidirectional recurrent neural networks
M. Schuster and K. K. Paliwal · 1997
Earlier work this paper cites.
Direct and indirect effects
J. Pearl · 2001
Earlier work this paper cites.
Causation, Prediction, and Search
P. Spirtes, C. N. Glymour, and R. Scheines · 2001
Earlier work this paper cites.
Natural logic for textual inference
B. MacCartney and C. D. Manning · 2007
Earlier work this paper cites.
A brief history of natural logic
J. van Benthem · 2008
Earlier work this paper cites.
An extended model of natural logic
B. MacCartney and C. D. Manning · 2009
Earlier work this paper cites.
Recent progress on monotonicity
T. F. Icard and L. S. Moss · 2013
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
J. Springenberg, A. Dosovitskiy, T. Brox, and M. Riedmiller · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Earlier work this paper cites.
Causal inference in statistics, social, and biomedical sciences
G. W. Imbens and D. B. Rubin · 2015
Earlier work this paper cites.
Layer-wise relevance propagation for neural networks with local renormalization layers
A. Binder, G. Montavon, S. Bach, K. Müller, and W. Samek · 2016
Earlier work this paper cites.
Multi-level cause-effect systems
K. Chalupka, F. Eberhardt, and P. Perona · 2016
Cited alongside, same era.
Not just a black box: Learning important features through propagating activation differences
A. Shrikumar, P. Greenside, A. Shcherbina, and A. Kundaje · 2016
Cited alongside, same era.
Causal consistency of structural equation models
P. K. Rubenstein, S. Weichwald, S. Bongers, J. M. Mooij, D. Janzing, M. Grosse-Wentrup, and B. Schölkopf · 2017
Cited alongside, same era.
Axiomatic attribution for deep networks
M. Sundararajan, A. Taly, and Q. Yan · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. u. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Analysing the potential of seq-to-seq models for incremental interpretation in task-oriented dialogue
Posing fair generalization tasks for natural language inference
A. Geiger, I. Cases, L. Karttunen, and C. Potts · 2019
Later among the works it cites.
Designing and interpreting probes with control tasks
J. Hewitt and P. Liang · 2019
Later among the works it cites.
What does it mean to understand a neural network?, 2019
T. P. Lillicrap and K. P. Kording · 2019
Later among the works it cites.
BERT rediscovers the classical NLP pipeline
I. Tenney, D. Das, and E. Pavlick · 2019
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, and J. Brew · 2019
Later among the works it cites.
Approximate causal abstractions
S. Beckers, F. Eberhardt, and J. Y. Halpern · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Hupkes, S. Bouwmeester, and R. Fernández · 2018
Cited alongside, same era.
Dissecting contextual word embeddings: Architecture and representation
M. Peters, M. Neumann, L. Zettlemoyer, and W.-t. Yih · 2018
Cited alongside, same era.
Abstracting causal models
S. Beckers and J. Y. Halpern · 2019
Cited alongside, same era.
Neural network attributions: A causal perspective
A. Chattopadhyay, P. Manupriya, A. Sarkar, and V. N. Balasubramanian · 2019
Cited alongside, same era.
What does BERT look at? an analysis of BERT’s attention
K. Clark, U. Khandelwal, O. Levy, and C. D. Manning · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Cited alongside, same era.
Later among the works it cites.
Foundations of structural causal models with cycles and latent variables
S. Bongers, P. Forré, J. Peters, B. Schölkopf, and J. M. Mooij · 2020
Later among the works it cites.
Amnesic probing: Behavioral explanation with amnesic counterfactuals
Y. Elazar, S. Ravfogel, A. Jacovi, and Y. Goldberg · 2020
Later among the works it cites.
Neural natural language inference models partially embed theories of lexical entailment and negation
A. Geiger, K. Richardson, and C. Potts · 2020
Later among the works it cites.
Probing the probing paradigm: Does probing accuracy entail task relevance?, 2020
A. Ravichander, Y. Belinkov, and E. Hovy · 2020
Later among the works it cites.