Fetching the paper…
Reading the bibliography…
In this paper, we propose a comprehensive benchmark to investigate models' logical reasoning capabilities in complex real-life scenarios.
METGEN: A module-based entailment tree generation framework for answer explanation
Ruixin Hong, Hongming Zhang, Xintong Yu, and Changshui Zhang. 2022 · 1905
Earlier work this paper cites.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
A New Introduction to Modal Logic
George Edward Hughes, Max J Cresswell, and Mary Meyerhoff Cresswell. 1996 · 1996
Earlier work this paper cites.
Building a discourse-tagged corpus in the framework of rhetorical structure theory
Lynn Carlson, Daniel Marcu, and Mary Ellen Okurovsky. 2001 · 2001
Earlier work this paper cites.
The uses of argument
Stephen E Toulmin. 2003 · 2003
Earlier work this paper cites.
Annotating argument components and relations in persuasive essays
Christian Stab and Iryna Gurevych. 2014a · 2014
Earlier work this paper cites.
Identifying argumentative discourse structures in persuasive essays
Christian Stab and Iryna Gurevych. 2014b · 2014
Earlier work this paper cites.
Parsing argumentation structures in persuasive essays
Christian Stab and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Adafactor: Adaptive learning rates with sublinear memory cost
Noam Shazeer and Mitchell Stern. 2018 · 2018
Earlier work this paper cites.
Argument mining: A survey
John Lawrence and Chris Reed. 2019 · 2019
Earlier work this paper cites.
The penn discourse treebank 3.0 annotation manual
Bonnie Webber, Rashmi Prasad, Alan Lee, and Aravind Joshi. 2019 · 2019
Cited alongside, same era.
Transformers as soft reasoners over language
Peter Clark, Oyvind Tafjord, and Kyle Richardson. 2020 · 2020
Cited alongside, same era.
MuTual: A dataset for multi-turn dialogue reasoning
Leyang Cui, Yu Wu, Shujie Liu, Yue Zhang, and Ming Zhou. 2020 · 2020
Cited alongside, same era.
ERASER: A benchmark to evaluate rationalized NLP models
Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric Lehman, Caiming Xiong, Richard Socher, and Byron C. Wallace. 2020 · 2020
Cited alongside, same era.
R4C: A benchmark for evaluating RC systems to get the right answer for the right reason
Naoya Inoue, Pontus Stenetorp, and Kentaro Inui. 2020 · 2020
Cited alongside, same era.
Learning to explain: Datasets and models for identifying valid reasoning chains in multihop question-answering
Modal Logic
James Garson. 2021 · 2021
Later among the works it cites.
DAGN: Discourse-aware graph network for logical reasoning
Yinya Huang, Meng Fang, Yu Cao, Liwei Wang, and Xiaodan Liang. 2021 · 2021
Later among the works it cites.
Predicting inductive biases of pre-trained models
Charles Lovering, Rohan Jha, Tal Linzen, and Ellie Pavlick. 2021 · 2021
Later among the works it cites.
Siru Ouyang, Zhuosheng Zhang, and Hai Zhao. 2021 · 2021
Later among the works it cites.
ExplaGraphs: An explanation graph generation task for structured commonsense reasoning
Swarnadeep Saha, Prateek Yadav, Lisa Bauer, and Mohit Bansal. 2021 · 2021
Later among the works it cites.
ProofWriter: Generating implications, proofs, and abductive statements over natural language
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Harsh Jhamtani and Peter Clark. 2020 · 2020
Cited alongside, same era.
LogiQA: A challenge dataset for machine reading comprehension with logical reasoning
Jian Liu, Leyang Cui, Hanmeng Liu, Dandan Huang, Yile Wang, and Yue Zhang. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
ReClor: A reading comprehension dataset requiring logical reasoning
Weihao Yu, Zihang Jiang, Yanfei Dong, and Jiashi Feng. 2020 · 2020
Cited alongside, same era.
Explaining answers with entailment trees
Bhavana Dalvi, Peter Jansen, Oyvind Tafjord, Zhengnan Xie, Hannah Smith, Leighanna Pipatanangkura, and Peter Clark. 2021 · 2021
Cited alongside, same era.
Oyvind Tafjord, Bhavana Dalvi, and Peter Clark. 2021 · 2021
Later among the works it cites.
A survey of discourse parsing
Jiaqi Li, Ming Liu, Bing Qin, and Ting Liu. 2022 · 2022
Closest in time.
Logic-driven context extension and data augmentation for logical reasoning of text
Siyuan Wang, Wanjun Zhong, Duyu Tang, Zhongyu Wei, Zhihao Fan, Daxin Jiang, Ming Zhou, and Nan Duan. 2022 · 2022
Closest in time.
Logiformer: A two-branch graph transformer network for interpretable logical reasoning
Fangzhi Xu, Jun Liu, Qika Lin, Yudai Pan, and Lingling Zhang. 2022 · 2022
Closest in time.