Fetching the paper…
Reading the bibliography…
Natural Language Inference (NLI) datasets contain annotation artefacts resulting in spurious correlations between the natural language utterances and their respective entailment classes.
On gradient descent ascent for nonconvex-concave minimax problems
Tianyi Lin, Chi Jin, and Michael I. Jordan. 2019 · 1906
Earlier work this paper cites.
On a test of whether one of two random variables is stochastically larger than the other
Henry B Mann and Donald R Whitney. 1947 · 1947
Earlier work this paper cites.
An Introduction to the Bootstrap
Bradley Efron and Robert Tibshirani. 1993 · 1993
Earlier work this paper cites.
Multiple hypothesis testing
Juliet Popper Shaffer. 1995 · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Sheng Zhang, Rachel Rudinger, Kevin Duh, and Benjamin Van Durme. 2017 · 2004
Earlier work this paper cites.
What syntax can contribute in the entailment task
Lucy Vanderwende and William B. Dolan. 2005 · 2005
Earlier work this paper cites.
Effectively using syntax for recognizing false entailment
Rion Snow, Lucy Vanderwende, and Arul Menezes. 2006 · 2006
Earlier work this paper cites.
Resolving complex cases of definite pronouns: The winograd schema challenge
Altaf Rahman and Vincent Ng. 2012 · 2012
Earlier work this paper cites.
A SICK cure for the evaluation of compositional distributional semantic models
Marco Marelli, Stefano Menini, Marco Baroni, Luisa Bentivogli, Raffaella Bernardi, and Roberto Zamparelli. 2014 · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian J. Goodfellow, and Rob Fergus. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Unsupervised domain adaptation by backpropagation
Yaroslav Ganin and Victor S. Lempitsky. 2015 · 2015
Earlier work this paper cites.
Framenet+: Fast paraphrastic tripling of framenet
Ellie Pavlick, Travis Wolfe, Pushpendre Rastogi, Chris Callison-Burch, Mark Dredze, and Benjamin Van Durme. 2015 · 2015
Earlier work this paper cites.
Semantic proto-roles
Drew Reisinger, Rachel Rudinger, Francis Ferraro, Craig Harman, Kyle Rawlins, and Benjamin Van Durme. 2015 · 2015
Earlier work this paper cites.
Deep Learning
Ian J. Goodfellow, Yoshua Bengio, and Aaron C. Courville. 2016 · 2016
Earlier work this paper cites.
The goldilocks principle: Reading children’s books with explicit memory representations
Felix Hill, Antoine Bordes, Sumit Chopra, and Jason Weston. 2016 · 2016
Earlier work this paper cites.
Answer-type prediction for visual question answering
Kushal Kafle and Christopher Kanan. 2016 · 2016
Cited alongside, same era.
A corpus and cloze evaluation for deeper understanding of commonsense stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James F. Allen. 2016 · 2016
Cited alongside, same era.
Most ”babies” are ”little” and most ”problems” are ”huge”: Compositional entailment in adjective-nouns
Ellie Pavlick and Chris Callison-Burch. 2016 · 2016
Cited alongside, same era.
Towards ai-complete question answering: A set of prerequisite toy tasks
Jason Weston, Antoine Bordes, Sumit Chopra, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
Yin and yang: Balancing and answering binary visual questions
Peng Zhang, Yash Goyal, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2016 · 2016
Cited alongside, same era.
Adversarial example generation with syntactically controlled paraphrase networks
Mohit Iyyer, John Wieting, Kevin Gimpel, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
How much reading does reading comprehension require? A critical investigation of popular benchmarks
Divyansh Kaushik and Zachary C. Lipton. 2018 · 2018
Later among the works it cites.
Scitail: A textual entailment dataset from science question answering
Tushar Khot, Ashish Sabharwal, and Peter Clark. 2018 · 2018
Later among the works it cites.
Adversarially regularising neural NLI models to integrate logical background knowledge
Pasquale Minervini and Sebastian Riedel. 2018 · 2018
Later among the works it cites.
Hypothesis only baselines in natural language inference
Adam Poliak, Jason Naradowsky, Aparajita Haldar, Rachel Rudinger, and Benjamin Van Durme. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zheng Cai, Lifu Tu, and Kevin Gimpel. 2017 · 2017
Cited alongside, same era.
Supervised learning of universal sentence representations from natural language inference data
Alexis Conneau, Douwe Kiela, Holger Schwenk, Loïc Barrault, and Antoine Bordes. 2017 · 2017
Cited alongside, same era.
Making the V in VQA matter: Elevating the role of image understanding in visual question answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2017 · 2017
Cited alongside, same era.
Deceiving google’s cloud video intelligence API built for summarizing videos
Hossein Hosseini, Baicen Xiao, and Radha Poovendran. 2017 · 2017
Cited alongside, same era.
Natural language inference from multiple premises
Alice Lai, Yonatan Bisk, and Julia Hockenmaier. 2017 · 2017
Cited alongside, same era.
The effect of different writing tasks on linguistic style: A case study of the ROC story cloze task
Roy Schwartz, Maarten Sap, Ioannis Konstas, Leila Zilles, Yejin Choi, and Noah A. Smith. 2017 · 2017
Cited alongside, same era.
Inference is everything: Recasting semantic resources into a unified evaluation framework
Aaron Steven White, Pushpendre Rastogi, Kevin Duh, and Benjamin Van Durme. 2017 · 2017
Cited alongside, same era.
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Later among the works it cites.
Neural network models for natural language inference fail to capture the semantics of inference
Aarne Talman and Stergios Chatzikyriakidis. 2018 · 2018
Later among the works it cites.
Performance impact caused by hidden bias of training data for recognizing textual entailment
Masatoshi Tsuchiya. 2018 · 2018
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
Robust machine comprehension models via adversarial training
Yicheng Wang and Mohit Bansal. 2018 · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel R. Bowman. 2018 · 2018
Later among the works it cites.
Don’t take the easy way out: Ensemble based methods for avoiding known dataset biases
Christopher Clark, Mark Yatskar, and Luke Zettlemoyer. 2019 · 2019
Later among the works it cites.
Unlearn dataset bias in natural language inference by fitting the residual
He He, Sheng Zha, and Haohan Wang. 2019 · 2019
Later among the works it cites.
Deep adversarial learning for NLP
William Yang Wang, Sameer Singh, and Jiwei Li. 2019 · 2019
Later among the works it cites.
End-to-end bias mitigation by modelling biases in corpora
Rabeeh Karimi Mahabadi, Yonatan Belinkov, and James Henderson. 2020 · 2020
Closest in time.
Adversarial examples for evaluating reading comprehension systems
Robin Jia and Percy Liang. 2017 · 2031
Closest in time.