Fetching the paper…
Reading the bibliography…
With the development and proliferation of large, complex, black-box models for solving many natural language processing (NLP) tasks, there is also an increasing necessity of methods to stress-test these models and provide some degree of interpretability or explainability.
V. I. Levenshtein et al. , “Binary codes capable of correcting deletions, insertions, and reversals,” in Soviet physics doklady , vol. 10, no. 8. Soviet Union, 1966, pp. 707–710
1966
Earlier work this paper cites.
B. MacCartney and C. D. Manning, “Modeling semantic containment and exclusion in natural language inference,” in Proceedings of the 22nd International Conference on Computational Linguistics (Coling 2008) , 2008, pp. 521–528
2008
Earlier work this paper cites.
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts, “Learning word vectors for sentiment analysis,” in Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies . Portland, Oregon, USA: Association for Computational Linguistics, June 2011, pp. 142–150. [Online]. Available: http://www.aclweb.org/anthology/P11-1015
2011
Earlier work this paper cites.
S. Bowman, G. Angeli, C. Potts, and C. D. Manning, “A large annotated corpus for learning natural language inference,” in Proceedings of EMNLP , 2015, pp. 632–642
2015
Earlier work this paper cites.
M. T. Ribeiro, S. Singh, and C. Guestrin, “” why should i trust you?” explaining the predictions of any classifier,” in Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining , 2016, pp. 1135–1144
2016
Earlier work this paper cites.
P. F. Christiano, J. Leike, T. Brown, M. Martic, S. Legg, and D. Amodei, “Deep reinforcement learning from human preferences,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
O.-M. Camburu, T. Rocktäschel, T. Lukasiewicz, and P. Blunsom, “e-snli: Natural language inference with natural language explanations,” Advances in Neural Information Processing Systems , vol. 31, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Wang, Y. Pruksachatkun, N. Nangia, A. Singh, J. Michael, F. Hill, O. Levy, and S. Bowman, “Superglue: A stickier benchmark for general-purpose language understanding systems,” Advances in neural information processing systems , vol. 32, 2019
2019
Earlier work this paper cites.
L. Qin, A. Bosselut, A. Holtzman, C. Bhagavatula, E. Clark, and Y. Choi, “Counterfactual story reasoning and generation,” in Proceedings of EMNLP-IJCNLP , 2019, pp. 5043–5053
2019
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
C. Molnar, Interpretable machine learning . Lulu. com, 2020
2020
Earlier work this paper cites.
M. Gardner, Y. Artzi, V. Basmov, J. Berant, B. Bogin, S. Chen, P. Dasigi, D. Dua, Y. Elazar, A. Gottumukkala et al. , “Evaluating models’ local decision boundaries via contrast sets,” in Findings of the ACL: EMNLP 2020 , 2020, pp. 1307–1323
2020
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
P. Atanasova, J. G. Simonsen, C. Lioma, and I. Augenstein, “A diagnostic study of explainability techniques for text classification,” in Proceedings of EMNLP , 2020, pp. 3256–3274
2020
Cited alongside, same era.
D. Kaushik, E. Hovy, and Z. Lipton, “Learning the difference that makes a difference with counterfactually-augmented data,” in ICLR , 2020. [Online]. Available: https://openreview.net/forum?id=Sklgs0NFvr
2020
Cited alongside, same era.
D. Khashabi, T. Khot, and A. Sabharwal, “More bang for your buck: Natural perturbation for robust question answering,” in Proceedings of EMNLP , 2020, pp. 163–170
2020
Cited alongside, same era.
S. Garg and G. Ramakrishnan, “Bae: Bert-based adversarial examples for text classification,” in Proceedings of EMNLP , 2020, pp. 6174–6181
2020
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
G. Penedo, Q. Malartic, D. Hesslow, R. Cojocaru, A. Cappelli, H. Alobeidli, B. Pannier, E. Almazrouei, and J. Launay, “The refinedweb dataset for falcon llm: Outperforming curated corpora with web data, and web data only,” CoRR , 2023
2023
Later among the works it cites.
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
2021
Cited alongside, same era.
T. Wu, M. T. Ribeiro, J. Heer, and D. S. Weld, “Polyjuice: Generating counterfactuals for explaining, evaluating, and improving models,” in Proceedings of the 59th Annual Meeting of the ACL and the 11th IJCNLP (Volume 1: Long Papers) , 2021, pp. 6707–6723
2021
Cited alongside, same era.
N. Madaan, I. Padhi, N. Panwar, and D. Saha, “Generate your counterfactuals: Towards controlled counterfactual generation for text,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 35, no. 15, 2021, pp. 13 516–13 524
2021
Cited alongside, same era.
M. Robeer, F. Bex, and A. Feelders, “Generating realistic natural language counterfactuals,” in Findings of the Association for Computational Linguistics: EMNLP 2021 , 2021, pp. 3611–3625
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2022
Cited alongside, same era.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” Advances in Neural Information Processing Systems , vol. 35, pp. 27 730–27 744, 2022
2022
Cited alongside, same era.
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
F. Gilardi, M. Alizadeh, and M. Kubli, “Chatgpt outperforms crowd workers for text-annotation tasks,” Proceedings of the National Academy of Sciences , vol. 120, no. 30, p. e2305016120, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Ali, T. Abuhmed, S. El-Sappagh, K. Muhammad, J. M. Alonso-Moral, R. Confalonieri, R. Guidotti, J. Del Ser, N. Díaz-Rodríguez, and F. Herrera, “Explainable artificial intelligence (xai): What we know and what is left to attain trustworthy artificial intelligence,” Information Fusion , vol. 99, p. 101805, 2023
2023
Later among the works it cites.
A. Bhattacharjee, R. Moraffah, J. Garland, and H. Liu, “Towards llm-guided causal explainability for black-box text classifiers,” in AAAI 2024 Workshop on Responsible Language Models, Vancouver, BC, Canada , 2024
2024
Closest in time.
Y. Li, M. Xu, X. Miao, S. Zhou, and T. Qian, “Prompting large language models for counterfactual generation: An empirical study,” in Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) , 2024, pp. 13 201–13 221
2024
Closest in time.
2024
Closest in time.
H. W. Chung, L. Hou, S. Longpre, B. Zoph, Y. Tay, W. Fedus, Y. Li, X. Wang, M. Dehghani, S. Brahma et al. , “Scaling instruction-finetuned language models,” JMLR , vol. 25, no. 70, pp. 1–53, 2024
2024
Closest in time.
Q. Dong, L. Li, D. Dai, C. Zheng, J. Ma, R. Li, H. Xia, J. Xu, Z. Wu, B. Chang et al. , “A survey on in-context learning,” in Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , 2024, pp. 1107–1128
2024
Closest in time.
2024
Closest in time.
A. Bhattacharjee and H. Liu, “Fighting fire with fire: can chatgpt detect ai-generated text?” ACM SIGKDD Explorations Newsletter , vol. 25, no. 2, pp. 14–21, 2024
2024
Closest in time.