Fetching the paper…
Reading the bibliography…
As an indispensable ingredient of intelligence, commonsense reasoning is crucial for large language models (LLMs) in real-world scenarios.
Bleu: a Method for Automatic Evaluation of Machine Translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
ConceptNet—a practical commonsense reasoning tool-kit
Liu, H.; and Singh, P. 2004 · 2004
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments
Banerjee, S.; and Lavie, A. 2005 · 2005
Earlier work this paper cites.
Isanette: A Common and Common Sense Knowledge Base for Opinion Mining
Cambria, E.; Song, Y.; Wang, H.; and Hussain, A. 2011 · 2011
Earlier work this paper cites.
The winograd schema challenge
Levesque, H.; Davis, E.; and Morgenstern, L. 2012 · 2012
Earlier work this paper cites.
Interrater reliability: the kappa statistic
McHugh, M. L. 2012 · 2012
Earlier work this paper cites.
CIDEr: Consensus-Based Image Description Evaluation
Vedantam, R.; Lawrence Zitnick, C.; and Parikh, D. 2015 · 2015
Earlier work this paper cites.
Commonsense for generative multi-hop question answering tasks
Bauer, L.; Wang, Y.; and Bansal, M. 2018 · 2018
Earlier work this paper cites.
MultiWOZ - A Large-Scale Multi-Domain Wizard-of-Oz Dataset for Task-Oriented Dialogue Modelling
Budzianowski, P.; Wen, T.-H.; Tseng, B.-H.; Casanueva, I.; Ultes, S.; Ramadan, O.; and Gašić, M. 2018 · 2018
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Clark, P.; Cowhey, I.; Etzioni, O.; Khot, T.; Sabharwal, A.; Schoenick, C.; and Tafjord, O. 2018 · 2018
Earlier work this paper cites.
Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering
Mihaylov, T.; Clark, P.; Khot, T.; and Sabharwal, A. 2018 · 2018
Earlier work this paper cites.
KagNet: Knowledge-Aware Graph Networks for Commonsense Reasoning
Lin, B. Y.; Chen, X.; Chen, J.; and Ren, X. 2019 · 2019
Earlier work this paper cites.
Social IQa: Commonsense Reasoning about Social Interactions
Sap, M.; Rashkin, H.; Chen, D.; Le Bras, R.; and Choi, Y. 2019b · 2019
Earlier work this paper cites.
Recent Advances in Natural Language Inference: A Survey of Benchmarks, Resources, and Approaches
Storks, S.; Gao, Q.; and Chai, J. Y. 2019 · 2019
Earlier work this paper cites.
CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge
Talmor, A.; Herzig, J.; Lourie, N.; and Berant, J. 2019 · 2019
Earlier work this paper cites.
HellaSwag: Can a Machine Really Finish Your Sentence?
Zellers, R.; Holtzman, A.; Bisk, Y.; Farhadi, A.; and Choi, Y. 2019 · 2019
Earlier work this paper cites.
“Going on a vacation” takes longer than “Going for a walk”: A Study of Temporal Commonsense Understanding
Zhou, B.; Khashabi, D.; Ning, Q.; and Roth, D. 2019 · 2019
Cited alongside, same era.
Piqa: Reasoning about physical commonsense in natural language
Bisk, Y.; Zellers, R.; Gao, J.; and Choi, Y. 2020 · 2020
Cited alongside, same era.
(Comet-) atomic 2020: on symbolic and neural commonsense knowledge graphs
Hwang, J. D.; Bhagavatula, C.; Le Bras, R.; Da, J.; Sakaguchi, K.; Bosselut, A.; and Choi, Y. 2021 · 2020
Cited alongside, same era.
Qasc: A dataset for question answering via sentence composition
Khot, T.; Clark, P.; Guerquin, M.; Jansen, P.; and Sabharwal, A. 2020 · 2020
Cited alongside, same era.
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-Trained Language Models
Lin, B. Y.; Lee, S.; Khanna, R.; and Ren, X. 2020 · 2020
Cited alongside, same era.
Commonsense knowledge reasoning and generation with pre-trained language models: A survey
Bhargava, P.; and Ng, V. 2022 · 2022
Later among the works it cites.
CICERO: A Dataset for Contextualized Commonsense Inference in Dialogues
Ghosal, D.; Shen, S.; Majumder, N.; Mihalcea, R.; and Poria, S. 2022 · 2022
Later among the works it cites.
An empirical analysis of compute-optimal large language model training
Hoffmann, J.; Borgeaud, S.; Mensch, A.; Buchatskaya, E.; Cai, T.; Rutherford, E.; de Las Casas, D.; Hendricks, L. A.; Welbl, J.; and Clark, A. 2022 · 2022
Later among the works it cites.
LoRA: Low-Rank Adaptation of Large Language Models
Hu, E. J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W. 2022 · 2022
Later among the works it cites.
Generated Knowledge Prompting for Commonsense Reasoning
Liu, J.; Liu, A.; Lu, X.; Welleck, S.; West, P.; Le Bras, R.; Choi, Y.; and Hajishirzi, H. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Graph-Based Reasoning over Heterogeneous External Knowledge for Commonsense Question Answering
Lv, S.; Guo, D.; Xu, J.; Tang, D.; Duan, N.; Gong, M.; Shou, L.; Jiang, D.; Cao, G.; and Hu, S. 2020 · 2020
Cited alongside, same era.
RiSAWOZ: A Large-Scale Multi-Domain Wizard-of-Oz Dataset with Rich Semantic Annotations for Task-Oriented Dialogue Modeling
Quan, J.; Zhang, S.; Cao, Q.; Li, Z.; and Xiong, D. 2020 · 2020
Cited alongside, same era.
Connecting the Dots: A Knowledgeable Path Generator for Commonsense Question Answering
Wang, P.; Peng, N.; Ilievski, F.; Szekely, P.; and Ren, X. 2020 · 2020
Cited alongside, same era.
Evaluating commonsense in pre-trained language models
Zhou, X.; Zhang, Y.; Cui, L.; and Huang, D. 2020 · 2020
Cited alongside, same era.
CrossWOZ: A Large-Scale Chinese Cross-Domain Task-Oriented Dialogue Dataset
Zhu, Q.; Huang, K.; Zhang, Z.; Zhu, X.; and Huang, M. 2020 · 2020
Cited alongside, same era.
CIDER: Commonsense Inference for Dialogue Explanation and Reasoning
Ghosal, D.; Hong, P.; Shen, S.; Majumder, N.; Mihalcea, R.; and Poria, S. 2021 · 2021
Cited alongside, same era.
“I’m Not Mad”: Commonsense Implications of Negation and Contradiction
Jiang, L.; Bosselut, A.; Bhagavatula, C.; and Choi, Y. 2021 · 2021
Cited alongside, same era.
Muennighoff, N.; Wang, T.; Sutawika, L.; Roberts, A.; Biderman, S.; Scao, T. L.; Bari, M. S.; Shen, S.; Yong, Z. X.; Schoelkopf, H.; Tang, X.; Radev, D.; Aji, A. F.; Almubarak, K.; Albanie, S.; Alyafeai, Z.; Webson, A.; Raff, E.; and Raffel, C. 2022 · 2022
Later among the works it cites.
Bloom: A 176b-parameter open-access multilingual language model
Scao, T. L.; Fan, A.; Akiki, C.; Pavlick, E.; Ilić, S.; Hesslow, D.; Castagné, R.; Luccioni, A. S.; Yvon, F.; and Gallé, M. 2022 · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; and Zhou, D. 2022 · 2022
Later among the works it cites.
Symbolic Knowledge Distillation: from General Language Models to Commonsense Models
West, P.; Bhagavatula, C.; Hessel, J.; Hwang, J.; Jiang, L.; Le Bras, R.; Lu, X.; Welleck, S.; and Choi, Y. 2022 · 2022
Later among the works it cites.
Long time no see! open-domain conversation with long-term persona memory
Xu, X.; Gou, Z.; Wu, W.; Niu, Z.-Y.; Wu, H.; Wang, H.; and Wang, S. 2022 · 2022
Later among the works it cites.
ClarET: Pre-training a Correlation-Aware Context-To-Event Transformer for Event-Centric Generation and Classification
Zhou, Y.; Shen, T.; Geng, X.; Long, G.; and Jiang, D. 2022 · 2022
Later among the works it cites.
Bang, Y.; Cahyawijaya, S.; Lee, N.; Dai, W.; Su, D.; Wilie, B.; Lovenia, H.; Ji, Z.; Yu, T.; Chung, W.; et al. 2023 · 2023
Closest in time.
Bian, N.; Han, X.; Sun, L.; Lin, H.; Lu, Y.; and He, B. 2023 · 2023
Closest in time.
Efficient and effective text encoding for chinese llama and alpaca
Cui, Y.; Yang, Z.; and Yao, X. 2023 · 2023
Closest in time.
Evaluating large language models: A comprehensive survey
Guo, Z.; Jin, R.; Liu, C.; Huang, Y.; Shi, D.; Yu, L.; Liu, Y.; Li, J.; Xiong, B.; Xiong, D.; et al. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; and Azhar, F. 2023 · 2023
Closest in time.