Fetching the paper…
Reading the bibliography…
While large pre-trained language models (PLM) have shown their great skills at solving discriminative tasks, a significant gap remains when compared with humans for explanation-related tasks.
Language models are few-shot learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Zhang, T.; Kishore, V.; Wu, F.; Weinberger, K. Q.; and Artzi, Y. 2019 · 1904
Earlier work this paper cites.
Explain yourself! leveraging language models for commonsense reasoning
Rajani, N. F.; McCann, B.; Xiong, C.; and Socher, R. 2019 · 1906
Earlier work this paper cites.
Does it make sense? and why? a pilot study for sense making and explanation
Wang, C.; Liang, S.; Zhang, Y.; Li, X.; and Gao, T. 2019 · 1906
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Probing Neural Network Comprehension of Natural Language Arguments
Niven, T.; and Kao, H.-Y. 2019 · 1907
Earlier work this paper cites.
Geva, M.; Goldberg, Y.; and Berant, J. 2019 · 1908
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Reimers, N.; and Gurevych, I. 2019 · 1908
Earlier work this paper cites.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Zhao, W.; Peyrard, M.; Liu, F.; Gao, Y.; Meyer, C. M.; and Eger, S. 2019 · 1909
Earlier work this paper cites.
Contrastive explanation
Lipton, P. 1990 · 1990
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
Open mind common sense: Knowledge acquisition from the general public
Singh, P.; Lin, T.; Mueller, E. T.; Lim, G.; Perkins, T.; and Zhu, W. L. 2002 · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
Kalm at semeval-2020 task 4: Knowledge-aware language models for comprehension and generation
Wan, J.; and Huang, X. 2020 · 2005
Earlier work this paper cites.
BUT-FIT at SemEval-2020 Task 4: Multilingual Commonsense
Jon, J.; Fajcik, M.; Docekal, M.; and Smrz, P. 2020 · 2008
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Robertson, S.; Zaragoza, H.; et al. 2009 · 2009
Earlier work this paper cites.
A survey of the state of explainable AI for natural language processing
Danilevsky, M.; Qian, K.; Aharonov, R.; Katsis, Y.; Kawas, B.; and Sen, P. 2020 · 2010
Cited alongside, same era.
Explaining NLP models via minimal contrastive editing (MiCE)
Ross, A.; Marasović, A.; and Peters, M. E. 2020 · 2012
Cited alongside, same era.
Generating visual explanations
Hendricks, L. A.; Akata, Z.; Rohrbach, M.; Donahue, J.; Schiele, B.; and Darrell, T. 2016 · 2016
Cited alongside, same era.
Rationalizing neural predictions
Lei, T.; Barzilay, R.; and Jaakkola, T. 2016 · 2016
Cited alongside, same era.
Why we need new evaluation metrics for NLG
Novikova, J.; Dušek, O.; Curry, A. C.; and Rieser, V. 2017 · 2017
Reconstructing implicit knowledge with language models
Becker, M.; Liang, S.; and Frank, A. 2021 · 2021
Later among the works it cites.
Unsupervised Editing for Counterfactual Stories
Chen, J.; Gan, C.; Cheng, S.; Zhou, H.; Xiao, Y.; and Li, L. 2021 · 2021
Later among the works it cites.
Training verifiers to solve math word problems
Cobbe, K.; Kosaraju, V.; Bavarian, M.; Hilton, J.; Nakano, R.; Hesse, C.; and Schulman, J. 2021 · 2021
Later among the works it cites.
On Commonsense Cues in BERT for Solving Commonsense Tasks
Cui, L.; Cheng, S.; Wu, Y.; and Zhang, Y. 2021 · 2021
Later among the works it cites.
Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity
Lu, Y.; Bartolo, M.; Moore, A.; Riedel, S.; and Stenetorp, P. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
e-snli: Natural language inference with natural language explanations
Camburu, O.-M.; Rocktäschel, T.; Lukasiewicz, T.; and Blunsom, P. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2018 · 2018
Cited alongside, same era.
Multimodal explanations: Justifying decisions and pointing to the evidence
Park, D. H.; Hendricks, L. A.; Akata, Z.; Rohrbach, A.; Schiele, B.; Darrell, T.; and Rohrbach, M. 2018 · 2018
Cited alongside, same era.
Explainable AI planning (XAIP): overview and the case of contrastive explanation
Hoffmann, J.; and Magazzeni, D. 2019 · 2019
Cited alongside, same era.
Cgmh: Constrained sentence generation by metropolis-hastings sampling
Miao, N.; Zhou, H.; Mou, L.; Yan, R.; and Li, L. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford, A.; Wu, J.; Child, R.; Luan, D.; Amodei, D.; Sutskever, I.; et al. 2019 · 2019
Cited alongside, same era.
JUSTers at SemEval-2020 Task 4: Evaluating Transformer Models against Commonsense Validation and Explanation
Fadel, A.; Al-Ayyoub, M.; and Cambria, E. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Prompting contrastive explanations for commonsense reasoning tasks
Paranjape, B.; Michael, J.; Ghazvininejad, M.; Zettlemoyer, L.; and Hajishirzi, H. 2021 · 2021
Later among the works it cites.
Multitask prompted training enables zero-shot task generalization
Sanh, V.; Webson, A.; Raffel, C.; Bach, S. H.; Sutawika, L.; Alyafeai, Z.; Chaffin, A.; Stiegler, A.; Scao, T. L.; Raja, A.; et al. 2021 · 2021
Later among the works it cites.
Reframing human-ai collaboration for generating free-text explanations
Wiegreffe, S.; Hessel, J.; Swayamdipta, S.; Riedl, M.; and Choi, Y. 2021 · 2021
Later among the works it cites.
Teach me to explain: A review of datasets for explainable natural language processing
Wiegreffe, S.; and Marasovic, A. 2021 · 2021
Later among the works it cites.
Calibrate before use: Improving few-shot performance of language models
Zhao, Z.; Wallace, E.; Feng, S.; Klein, D.; and Singh, S. 2021 · 2021
Later among the works it cites.
Selection-Inference: Exploiting Large Language Models for Interpretable Logical Reasoning
Creswell, A.; Shanahan, M.; and Higgins, I. 2022 · 2022
Closest in time.
Self-consistency improves chain of thought reasoning in language models
Wang, X.; Wei, J.; Schuurmans, D.; Le, Q.; Chi, E.; and Zhou, D. 2022 · 2022
Closest in time.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Chi, E.; Le, Q.; and Zhou, D. 2022 · 2022
Closest in time.
Star: Bootstrapping reasoning with reasoning
Zelikman, E.; Wu, Y.; and Goodman, N. D. 2022 · 2022
Closest in time.
Opt: Open pre-trained transformer language models
Zhang, S.; Roller, S.; Goyal, N.; Artetxe, M.; Chen, M.; Chen, S.; Dewan, C.; Diab, M.; Li, X.; Lin, X. V.; et al. 2022 · 2022
Closest in time.
Describing Differences between Text Distributions with Natural Language
Zhong, R.; Snell, C.; Klein, D.; and Steinhardt, J. 2022 · 2022
Closest in time.