Event2Mind: Commonsense inference on events, intents, and reactions
Hannah Rashkin, Maarten Sap, Emily Allaway, Noah A. Smith, and Yejin Choi. 2018 · 2018
Later among the works it cites.
The role of veridicality and factivity in clause selection
Aaron Steven White and Kyle Rawlins. 2018 · 2018
Later among the works it cites.
Lexicosyntactic inference in neural models
Aaron Steven White, Rachel Rudinger, Kyle Rawlins, and Benjamin Van Durme. 2018 · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Later among the works it cites.
SWAG: A large-scale adversarial dataset for grounded commonsense inference
Rowan Zellers, Yonatan Bisk, Roy Schwartz, and Yejin Choi. 2018 · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Do you know that florence is packed with visitors? evaluating state-of-the-art models of speaker commitment
Nanjiang Jiang and Marie-Catherine de Marneffe. 2019 · 2019
Later among the works it cites.
Probing what different NLP tasks teach machines about function word comprehension
Najoung Kim, Roma Patel, Adam Poliak, Patrick Xia, Alex Wang, Tom McCoy, Ian Tenney, Alexis Ross, Tal Linzen, Benjamin Van Durme, Samuel R. Bowman, and Ellie Pavlick. 2019 · 2019
Later among the works it cites.
The CommitmentBank: Investigating projection in naturally occurring discourse
Marie-Catherine de Marneffe, Mandy Simons, and Judith Tonhauser. 2019 · 2019
Later among the works it cites.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
Tom McCoy, Ellie Pavlick, and Tal Linzen. 2019 · 2019
Later among the works it cites.
Analyzing compositionality-sensitivity of NLI models
Yixin Nie, Yicheng Wang, and Mohit Bansal. 2019 · 2019
Later among the works it cites.
SherLIiC: A typed event-focused lexical inference benchmark for evaluating natural language inference
Martin Schmitt and Hinrich Schütze. 2019 · 2019
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amapreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R Bowman. 2019 · 2019
Later among the works it cites.
Dialogue natural language inference
Sean Welleck, Jason Weston, Arthur Szlam, and Kyunghyun Cho. 2019 · 2019
Later among the works it cites.
Can neural networks understand monotonicity reasoning?
Hitomi Yanaka, Koji Mineshima, Daisuke Bekki, Kentaro Inui, Satoshi Sekine, Lasha Abzianidze, and Johan Bos. 2019a · 2019
Later among the works it cites.
HELP: A dataset for identifying shortcomings of neural models in monotonicity reasoning
Hitomi Yanaka, Koji Mineshima, Daisuke Bekki, Kentaro Inui, Satoshi Sekine, Lasha Abzianidze, and Johan Bos. 2019b · 2019
Later among the works it cites.
HellaSwag: Can a machine really finish your sentence?
Rowan Zellers, Ari Holtzman, Yonatan Bisk, Ali Farhadi, and Yejin Choi. 2019 · 2019
Later among the works it cites.
Abductive commonsense reasoning
Chandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya, Keisuke Sakaguchi, Ari Holtzman, Hannah Rashkin, Doug Downey, Scott Wen-tau Yih, and Yejin Choi. 2020 · 2020
Closest in time.
Probing natural language inference models through semantic fragments
Kyle Richardson, Hai Hu, Lawrence S Moss, and Ashish Sabharwal. 2020 · 2020
Closest in time.
Harnessing the richness of the linguistic signal in predicting pragmatic inferences
Sebastian Schuster, Yuxing Chen, and Judith Degen. 2020 · 2020
Closest in time.