Fetching the paper…
Reading the bibliography…
Explaining the predictions of AI models is paramount in safety-critical applications, such as in legal or medical domains.
Legal Judgment Prediction via Multi-perspective Bi-feedback Network
Yang, W.; Jia, W.; Zhou, X.; and Luo, Y. 2019 · 1905
Earlier work this paper cites.
TinyBERT: Distilling BERT for Natural Language Understanding
Jiao, X.; Yin, Y.; Shang, L.; Jiang, X.; Chen, X.; Li, L.; Wang, F.; and Liu, Q. 2019 · 1909
Earlier work this paper cites.
Your Classifier is Secretly an Energy Based Model and You Should Treat It Like One
Grathwohl, W.; Wang, K.-C.; Jacobsen, J.-H.; Duvenaud, D.; Norouzi, M.; and Swersky, K. 2019 · 1912
Earlier work this paper cites.
Simple Statistical Gradient-following Algorithms for Connectionist Reinforcement Learning
Williams, R. J. 1992 · 1992
Earlier work this paper cites.
Long Short-Term Memory
Hochreiter, S.; and Schmidhuber, J. 1997 · 1997
Earlier work this paper cites.
The Information Bottleneck Method
Tishby, N.; Pereira, F. C.; and Bialek, W. 2000 · 2000
Earlier work this paper cites.
Content Analysis: An Introduction to Its Methodology Thousand Oaks
Krippendorff, K. 2004 · 2004
Earlier work this paper cites.
Deep Neural Networks for Acoustic Modeling in Speech Recognition
Hinton, G.; Deng, L.; Yu, D.; Dahl, G.; Mohamed, A.-r.; Jaitly, N.; Senior, A.; Vanhoucke, V.; Nguyen, P.; Kingsbury, B.; and Sainath, T. 2012 · 2012
Earlier work this paper cites.
Learning Attitudes and Attributes From Multi-aspect Reviews
McAuley, J.; Leskovec, J.; and Jurafsky, D. 2012 · 2012
Earlier work this paper cites.
Auto-encoding Variational Bayes
Kingma, D. P.; and Welling, M. 2013 · 2013
Earlier work this paper cites.
Distributed Representations of Words and Phrases and Their Compositionality
Mikolov, T.; Sutskever, I.; Chen, K.; Corrado, G. S.; and Dean, J. 2013 · 2013
Earlier work this paper cites.
Generative Adversarial Nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Sutskever, I.; Vinyals, O.; and Le, Q. V. 2014 · 2014
Earlier work this paper cites.
Distilling the Knowledge in a Neural Network
Hinton, G.; Vinyals, O.; and Dean, J. 2015 · 2015
Earlier work this paper cites.
Deep Learning and The Information Bottleneck Principle
Tishby, N.; and Zaslavsky, N. 2015 · 2015
Earlier work this paper cites.
InfoGAN: Interpretable Representation Learning by Information Maximizing Generative Adversarial Nets
Chen, X.; Duan, Y.; Houthooft, R.; Schulman, J.; Sutskever, I.; and Abbeel, P. 2016 · 2016
Cited alongside, same era.
Generating Visual Explanations
Hendricks, L. A.; Akata, Z.; Rohrbach, M.; Donahue, J.; Schiele, B.; and Darrell, T. 2016 · 2016
Cited alongside, same era.
Categorical Reparameterization with Gumbel-softmax
Jang, E.; Gu, S.; and Poole, B. 2016 · 2016
Cited alongside, same era.
Rationalizing Neural Predictions
Lei, T.; Barzilay, R.; and Jaakkola, T. 2016 · 2016
Cited alongside, same era.
Reading and Thinking: Re-read LSTM Unit for Textual Entailment Recognition
Sha, L.; Chang, B.; Sui, Z.; and Li, S. 2016 · 2016
Cited alongside, same era.
Imagenet Classification with Deep Convolutional Neural Networks
Table-to-text Generation by Structure-aware Seq2seq Learning
Liu, T.; Wang, K.; Sha, L.; Chang, B.; and Sui, Z. 2018 · 2018
Later among the works it cites.
Multimodal Explanations: Justifying Decisions and Pointing to the Evidence
Park, D. H.; Hendricks, L. A.; Akata, Z.; Rohrbach, A.; Schiele, B.; Darrell, T.; and Rohrbach, M. 2018 · 2018
Later among the works it cites.
INVASE: Instance-wise Variable Selection Using Neural Networks
Yoon, J.; Jordon, J.; and van der Schaar, M. 2018 · 2018
Later among the works it cites.
Top-down Neural Attention by Excitation Backprop
Zhang, J.; Bargal, S. A.; Lin, Z.; Brandt, J.; Shen, X.; and Sclaroff, S. 2018 · 2018
Later among the works it cites.
Legal Judgment Prediction via Topological Learning
Zhong, H.; Guo, Z.; Tu, C.; Xiao, C.; Liu, Z.; and Sun, M. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Krizhevsky, A.; Sutskever, I.; and Hinton, G. E. 2017 · 2017
Cited alongside, same era.
Learning to Predict Charges for Criminal Cases with Legal Basis
Luo, B.; Feng, Y.; Xu, J.; Zhang, X.; and Zhao, D. 2017 · 2017
Cited alongside, same era.
Dynamic Routing Between Capsules
Sabour, S.; Frosst, N.; and Hinton, G. E. 2017 · 2017
Cited alongside, same era.
Seqgan: Sequence Generative Adversarial Nets with Policy Gradient
Yu, L.; Zhang, W.; Wang, J.; and Yu, Y. 2017 · 2017
Cited alongside, same era.
e-SNLI: Natural Language Inference with Natural Language Explanations
Camburu, O.; Rocktäschel, T.; Lukasiewicz, T.; and Blunsom, P. 2018 · 2018
Cited alongside, same era.
Extractive Adversarial Networks: High-Recall Explanations for Identifying Personal Attacks in Social Media Posts
Carton, S.; Mei, Q.; and Resnick, P. 2018 · 2018
Cited alongside, same era.
Learning to Explain: An Information-Theoretic Perspective on Model Interpretation
Chen, J.; Song, L.; Wainwright, M.; and Jordan, M. 2018 · 2018
Cited alongside, same era.
Bastings, J.; Aziz, W.; and Titov, I. 2019 · 2019
Later among the works it cites.
A Game Theoretic Approach to Class-wise Selective Rationalization
Chang, S.; Zhang, Y.; Yu, M.; and Jaakkola, T. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Later among the works it cites.
Specializing Word Embeddings (for Parsing) by Information Bottleneck
Li, X. L.; and Eisner, J. 2019 · 2019
Later among the works it cites.
Rethinking Cooperative Rationalization: Introspective Extraction and Complement Control
Yu, M.; Chang, S.; Zhang, Y.; and Jaakkola, T. 2019 · 2019
Later among the works it cites.
Gradient-guided Unsupervised Lexically Constrained Text Generation
Sha, L. 2020 · 2020
Closest in time.
Estimate Minimum Operation Steps via Memory-based Recurrent Calculation Network
Sha, L.; Shi, C.; Chen, Q.; Zhang, L.; and Wang, H. 2020 · 2020
Closest in time.
Multi-type Disentanglement without Adversarial Training
Sha, L.; and Lukasiewicz, T. 2021 · 2021
Closest in time.
A Progressive Learning Approach to Chinese SRL Using Heterogeneous Data
Xia, Q.; Sha, L.; Chang, B.; and Sui, Z. 2017 · 2077
Closest in time.