Fetching the paper…
Reading the bibliography…
A machine learning system can score well on a given test set by relying on heuristics that are effective for frequent example types but break down in more challenging cases.
Assessing BERT’s syntactic abilities
Yoav Goldberg. 2019 · 1901
Earlier work this paper cites.
Analyzing the behavior of visual question answering models
Aishwarya Agrawal, Dhruv Batra, and Devi Parikh. 2016 · 1960
Earlier work this paper cites.
The cognitive basis for linguistic structures
Thomas G. Bever. 1970 · 1970
Earlier work this paper cites.
Making and correcting errors during sentence comprehension: Eye movements in the analysis of structurally ambiguous sentences
Lyn Frazier and Keith Rayner. 1982 · 1982
Earlier work this paper cites.
Behavior analysis of NLI models: Uncovering the influence of three factors on robustness
Ivan Sanchez, Jeff Mitchell, and Sebastian Riedel. 2018 · 1985
Earlier work this paper cites.
Thematic roles assigned along the garden path linger
Kiel Christianson, Andrew Hollingworth, John F Halliwell, and Fernanda Ferreira. 2001 · 2001
Earlier work this paper cites.
Entailment, intensionality and text understanding
Cleo Condoravdi, Dick Crouch, Valeria de Paiva, Reinhard Stolle, and Daniel G. Bobrow. 2003 · 2003
Earlier work this paper cites.
Accurate unlexicalized parsing
Dan Klein and Christopher D. Manning. 2003 · 2003
Earlier work this paper cites.
Effects of merely local syntactic coherence on sentence processing
Whitney Tabor, Bruno Galantucci, and Daniel Richardson. 2004 · 2004
Earlier work this paper cites.
The PASCAL Recognising Textual Entailment Challenge
Ido Dagan, Oren Glickman, and Bernardo Magnini. 2006 · 2006
Earlier work this paper cites.
Natural language inference
Bill MacCartney and Christopher D Manning. 2009 · 2009
Earlier work this paper cites.
Syntactic/semantic structures for textual entailment recognition
Yashar Mehdad, Alessandro Moschitti, and Fabio Massimo Zanzotto. 2010 · 2010
Earlier work this paper cites.
Cambridge: Parser evaluation using textual entailment by grammatical relation comparison
Laura Rimell and Stephen Clark. 2010 · 2010
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
A fast unified model for parsing and sentence understanding
Samuel R. Bowman, Jon Gauthier, Abhinav Rastogi, Raghav Gupta, Christopher D. Manning, and Christopher Potts. 2016 · 2016
Earlier work this paper cites.
Assessing the ability of LSTMs to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Ankur Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Cited alongside, same era.
Tense manages to predict implicative behavior in verbs
Ellie Pavlick and Chris Callison-Burch. 2016 · 2016
Cited alongside, same era.
Fine-grained analysis of sentence embeddings using auxiliary prediction tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2017 · 2017
Cited alongside, same era.
Enhanced LSTM for natural language inference
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling, Si Wei, Hui Jiang, and Diana Inkpen. 2017 · 2017
Cited alongside, same era.
AllenNLP: A Deep Semantic Natural Language Processing Platform
Matt Gardner, Joel Grus, Mark Neumann, Oyvind Tafjord, Pradeep Dasigi, Nelson F. Liu, Matthew Peters, Michael Schmitz, and Luke S. Zettlemoyer. 2017 · 2017
Targeted syntactic evaluation of language models
Rebecca Marvin and Tal Linzen. 2018 · 2018
Later among the works it cites.
Stress test evaluation for natural language inference
Aakanksha Naik, Abhilasha Ravichander, Norman Sadeh, Carolyn Rose, and Graham Neubig. 2018 · 2018
Later among the works it cites.
Analyzing compositionality-sensitivity of NLI models
Yixin Nie, Yicheng Wang, and Mohit Bansal. 2018 · 2018
Later among the works it cites.
Collecting diverse natural language inference problems for sentence representation evaluation
Adam Poliak, Aparajita Haldar, Rachel Rudinger, J. Edward Hu, Ellie Pavlick, Aaron Steven White, and Benjamin Van Durme. 2018a · 2018
Later among the works it cites.
Neural models of factuality
Rachel Rudinger, Aaron Steven White, and Benjamin Van Durme. 2018 · 2018
Later among the works it cites.
Visual concepts and compositional voting
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Inference is everything: Recasting semantic resources into a unified evaluation framework
Aaron Steven White, Pushpendre Rastogi, Kevin Duh, and Benjamin Van Durme. 2017 · 2017
Cited alongside, same era.
What you can cram into a single vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, Germán Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Evaluating compositionality in sentence embeddings
Ishita Dasgupta, Demi Guo, Andreas Stuhlmüller, Samuel J. Gershman, and Noah D. Goodman. 2018 · 2018
Cited alongside, same era.
Assessing composition in sentence vector representations
Allyson Ettinger, Ahmed Elgohary, Colin Phillips, and Philip Resnik. 2018 · 2018
Cited alongside, same era.
Stress-testing neural models of natural language inference with multiply-quantified sentences
Atticus Geiger, Ignacio Cases, Lauri Karttunen, and Christopher Potts. 2018 · 2018
Cited alongside, same era.
Breaking NLI Systems with Sentences that Require Simple Lexical Inferences
Max Glockner, Vered Shwartz, and Yoav Goldberg. 2018 · 2018
Cited alongside, same era.
Jianyu Wang, Zhishuai Zhang, Cihang Xie, Yuyin Zhou, Vittal Premachandran, Jun Zhu, Lingxi Xie, and Alan Yuille. 2018 · 2018
Later among the works it cites.
The fine line between linguistic generalization and failure in seq2seq-attention models
Noah Weber, Leena Shekhar, and Niranjan Balasubramanian. 2018 · 2018
Later among the works it cites.
The role of veridicality and factivity in clause selection
Aaron Steven White and Kyle Rawlins. 2018 · 2018
Later among the works it cites.
Lexicosyntactic inference in neural models
Aaron Steven White, Rachel Rudinger, Kyle Rawlins, and Benjamin Van Durme. 2018 · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018b · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.
Non-entailed subsequences as a challenge for natural language inference
R. Thomas McCoy and Tal Linzen. 2019 · 2019
Closest in time.
RNNs implicitly implement tensor-product representations
R. Thomas McCoy, Tal Linzen, Ewan Dunbar, and Paul Smolensky. 2019 · 2019
Closest in time.
Human vs. muppet: A conservative estimate of human performance on the GLUE benchmark
Nikita Nangia and Samuel R. Bowman. 2019 · 2019
Closest in time.
Revisiting the poverty of the stimulus: Hierarchical generalization without a hierarchical bias in recurrent neural networks
R. Thomas McCoy, Robert Frank, and Tal Linzen. 2018 · 2098
Closest in time.