Roberta: A robustly optimized bert pretraining approach
Original
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Learning What Makes a Difference from Counterfactual Examples and Gradient Supervision
Original
Teney, D.; Abbasnedjad, E.; and van den Hengel, A. 2020 · 2004
Earlier work this paper cites.
Evaluating Explanations: How much do explanations from the teacher aid students?
Original
Pruthi, D.; Dhingra, B.; Soares, L. B.; Collins, M.; Lipton, Z. C.; Neubig, G.; and Cohen, W. W. 2020 · 2012
Earlier work this paper cites.
Self-Explaining Structures Improve NLP Models
Original
Sun, Z.; Fan, C.; Han, Q.; Sun, X.; Meng, Y.; Wu, F.; and Li, J. 2020 · 2012
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Bowman, S. R.; Angeli, G.; Potts, C.; and Manning, C. D. 2015 · 2015
Earlier work this paper cites.
Neural Architectures for Named Entity Recognition
Lample, G.; Ballesteros, M.; Subramanian, S.; Kawakami, K.; and Dyer, C. 2016 · 2016
Earlier work this paper cites.
“Why Should I Trust You?”: Explaining the Predictions of Any Classifier
Ribeiro, M.; Singh, S.; and Guestrin, C. 2016 · 2016
Earlier work this paper cites.
Learning with Latent Language
Andreas, J.; Klein, D.; and Levine, S. 2018 · 2018
Earlier work this paper cites.
e-SNLI: Natural Language Inference with Natural Language Explanations
Camburu, O.-M.; Rocktäschel, T.; Lukasiewicz, T.; and Blunsom, P. 2018 · 2018
Earlier work this paper cites.
Annotation Artifacts in Natural Language Inference Data
Gururangan, S.; Swayamdipta, S.; Levy, O.; Schwartz, R.; Bowman, S.; and Smith, N. A. 2018 · 2018
Earlier work this paper cites.
Adversarially Regularising Neural NLI Models to Integrate Logical Background Knowledge
Minervini, P.; and Riedel, S. 2018 · 2018
Earlier work this paper cites.
Hypothesis Only Baselines in Natural Language Inference
Poliak, A.; Naradowsky, J.; Haldar, A.; Rudinger, R.; and Van Durme, B. 2018 · 2018
Earlier work this paper cites.
Zero-Shot Sequence Labeling: Transferring Knowledge from Sentences to Tokens
Rei, M.; and Søgaard, A. 2018 · 2018
Earlier work this paper cites.
Performance Impact Caused by Hidden Bias of Training Data for Recognizing Textual Entailment
Tsuchiya, M. 2018 · 2018
Earlier work this paper cites.
A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference
Williams, A.; Nangia, N.; and Bowman, S. 2018 · 2018
Earlier work this paper cites.
Don’t Take the Easy Way Out: Ensemble Based Methods for Avoiding Known Dataset Biases
Clark, C.; Yatskar, M.; and Zettlemoyer, L. 2019 · 2019
Earlier work this paper cites.