Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated good performance in many reasoning tasks, but they still struggle with some complicated reasoning tasks including logical reasoning.
Help: A dataset for identifying shortcomings of neural models in monotonicity reasoning
Hitomi Yanaka, Koji Mineshima, Daisuke Bekki, Kentaro Inui, Satoshi Sekine, Lasha Abzianidze, and Johan Bos. 2019 · 1904
Earlier work this paper cites.
Logics and Languages
Maxwell John Cresswell. 1973 · 1973
Earlier work this paper cites.
Logic for problem solving
Robert Kowalski. 1974 · 1974
Earlier work this paper cites.
Logical reasoning in natural language: It is all about knowledge
Lucja Iwańska. 1993 · 1993
Earlier work this paper cites.
A Concise Introduction to Logic
Patrick J. Hurley. 2000 · 2000
Earlier work this paper cites.
Reclor: A reading comprehension dataset requiring logical reasoning
Weihao Yu, Zihang Jiang, Yanfei Dong, and Jiashi Feng. 2020 · 2002
Earlier work this paper cites.
Dare: Data augmented relation extraction with gpt-2
Yannis Papanikolaou and Andrea Pierleoni. 2020 · 2004
Earlier work this paper cites.
On sophistical refutations
Aristotle. 2006 · 2006
Earlier work this paper cites.
Logiqa: A challenge dataset for machine reading comprehension with logical reasoning
Jian Liu, Leyang Cui, Hanmeng Liu, Dandan Huang, Yile Wang, and Yue Zhang. 2020 · 2007
Earlier work this paper cites.
Fallacies and argument appraisal
Christopher W Tindale. 2007 · 2007
Earlier work this paper cites.
Taxinli: Taking a ride up the nlu hill
Pratik Joshi, Somak Aditya, Aalok Sathe, and Monojit Choudhury. 2020 · 2009
Earlier work this paper cites.
Semeval-2020 task 11: Detection of propaganda techniques in news articles
G Martino, Alberto Barrón-Cedeno, Henning Wachsmuth, Rostislav Petrov, and Preslav Nakov. 2020 · 2009
Earlier work this paper cites.
Case study research: What, why and how?
Peter Swanborn. 2010 · 2010
Earlier work this paper cites.
Logic and Philosophy: A Modern Introduction
Alan Hausman. 2012 · 2012
Earlier work this paper cites.
Synthetic text generation for sentiment analysis
Umar Maqsud. 2015 · 2015
Earlier work this paper cites.
Improving neural machine translation models with monolingual data
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2015 · 2015
Earlier work this paper cites.
Recognizing insufficiently supported arguments in argumentative essays
Christian Stab and Iryna Gurevych. 2017 · 2017
Cited alongside, same era.
Before name-calling: Dynamics and triggers of ad hominem fallacies in web argumentation
Ivan Habernal, Henning Wachsmuth, Iryna Gurevych, and Benno Stein. 2018 · 2018
Cited alongside, same era.
Language (technology) is power: A critical survey of “bias” in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2020
Cited alongside, same era.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Cited alongside, same era.
Natural language inference in context-investigating contextual reasoning over long texts
Hanmeng Liu, Leyang Cui, Jian Liu, and Yue Zhang. 2021 · 2021
Cited alongside, same era.
Can gpt-3 perform statutory reasoning?
Andrew Blair-Stanek, Nils Holzenberger, and Benjamin Van Durme. 2023 · 2023
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E Gonzalez, et al. 2023 · 2023
Later among the works it cites.
John Joon Young Chung, Ece Kamar, and Saleema Amershi. 2023 · 2023
Later among the works it cites.
Tinystories: How small can language models be and still speak coherent english?
Ronen Eldan and Yuanzhi Li. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
UTNLP at SemEval-2022 task 6: A comparative analysis of sarcasm detection using generative-based and mutation-based data augmentation
Amirhossein Abaskohi, Arash Rasouli, Tanin Zeraati, and Behnam Bahrak. 2022 · 2022
Cited alongside, same era.
Fallacious argument classification in political debates
Pierpaolo Goffredo, Shohreh Haddadan, Vorakit Vorakitphan, Elena Cabrio, and Serena Villata. 2022 · 2022
Cited alongside, same era.
Folio: Natural language reasoning with first-order logic
Simeng Han, Hailey Schoelkopf, Yilun Zhao, Zhenting Qi, Martin Riddell, Luke Benson, Lucy Sun, Ekaterina Zubova, Yujie Qiao, Matthew Burtell, et al. 2022 · 2022
Cited alongside, same era.
Towards reasoning in large language models: A survey
Jie Huang and Kevin Chen-Chuan Chang. 2022 · 2022
Cited alongside, same era.
Logical fallacy detection
Zhijing Jin, Abhinav Lalwani, Tejas Vaidhya, Xiaoyu Shen, Yiwen Ding, Zhiheng Lyu, Mrinmaya Sachan, Rada Mihalcea, and Bernhard Schoelkopf. 2022 · 2022
Cited alongside, same era.
Logicinference: A new dataset for teaching logical inference to seq2seq models
Santiago Ontanon, Joshua Ainslie, Vaclav Cvicek, and Zachary Fisher. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Cited alongside, same era.
Martin Josifoski, Marija Sakota, Maxime Peyrard, and Robert West. 2023 · 2023
Later among the works it cites.
Zhuoyan Li, Hangxiao Zhu, Zhuoran Lu, and Ming Yin. 2023 · 2023
Later among the works it cites.
Orca 2: Teaching small language models how to reason
Arindam Mitra, Luciano Del Corro, Shweti Mahajan, Andres Codas, Clarisse Simoes, Sahaj Agrawal, Xuxi Chen, Anastasia Razdaibiedina, Erik Jones, Kriti Aggarwal, Hamid Palangi, Guoqing Zheng, Corby Rosset, Hamed Khanpour, and Ahmed Awadallah. 2023 · 2023
Later among the works it cites.
Anders Giovanni Møller, Jacob Aarup Dalsgaard, Arianna Pera, and Luca Maria Aiello. 2023 · 2023
Later among the works it cites.
Gpt-4 technical report. arxiv 2303.08774
R OpenAI. 2023 · 2023
Later among the works it cites.
How susceptible are llms to logical fallacies?
Amirreza Payandeh, Dan Pluth, Jordan Hosier, Xuesu Xiao, and Vijay K Gurbani. 2023 · 2023
Later among the works it cites.
Case-based reasoning with language models for classification of logical fallacies
Zhivar Sourati, Filip Ilievski, Hông-Ân Sandlin, and Alain Mermoud. 2023 · 2023
Later among the works it cites.
Glore: Evaluating logical reasoning of large language models
Zhiyang Teng, Ruoxi Ning, Jian Liu, Qiji Zhou, Yue Zhang, et al. 2023 · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Later among the works it cites.
Nature language reasoning, a survey
Fei Yu, Hongbo Zhang, and Benyou Wang. 2023 · 2023
Later among the works it cites.
Improved logical reasoning of language models via differentiable symbolic programming
Hanlin Zhang, Jiani Huang, Ziyang Li, Mayur Naik, and Eric Xing. 2023 · 2023
Later among the works it cites.