Fetching the paper…
Reading the bibliography…
Commonsense AI has long been seen as a near impossible goal -- until recently.
SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems
Wang, A.; Pruksachatkun, Y.; Nangia, N.; Singh, A.; Michael, J.; Hill, F.; Levy, O.; and Bowman, S. R. 2019a · 1905
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019b · 1907
Earlier work this paper cites.
Schwartz, R.; Dodge, J.; Smith, N. A.; and Etzioni, O. 2019 · 1907
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Lan, Z.; Chen, M.; Goodman, S.; Gimpel, K.; Sharma, P.; and Soricut, R. 2019 · 1909
Earlier work this paper cites.
Programs with Common Sense
McCarthy, J. 1959 · 1959
Earlier work this paper cites.
Statistical Inference Under Order Restrictions: The Theory and Application of Isotonic Regression
Barlow, R.; Bartholomew, D.; Bremner, J.; and Brunk, H. 1972 · 1972
Earlier work this paper cites.
Understanding natural language
Winograd, T. 1972 · 1972
Earlier work this paper cites.
A Learning Algorithm for Continually Running Fully Recurrent Neural Networks
Williams, R. J.; and Zipser, D. 1989 · 1989
Earlier work this paper cites.
Direct Transfer of Learned Information Among Neural Networks
Pratt, L.; Mostow, J.; and Kamm, C. 1991 · 1991
Earlier work this paper cites.
Class-Based n
Brown, P. F.; Della Pietra, V. J.; deSouza, P. V.; Lai, J. C.; and Mercer, R. L. 1992 · 1992
Earlier work this paper cites.
Using qualitative models to guide inductive learning
Clark, P.; and Matwin, S. 1993 · 1993
Earlier work this paper cites.
Learning Many Related Tasks at the Same Time with Backpropagation
Caruana, R. 1995 · 1995
Earlier work this paper cites.
Scaling laws for neural language models
Kaplan, J.; McCandlish, S.; Henighan, T.; Brown, T. B.; Chess, B.; Child, R.; Gray, S.; Radford, A.; Wu, J.; and Amodei, D. 2020 · 2001
Earlier work this paper cites.
Adversarial Filters of Dataset Biases
Le Bras, R.; Swayamdipta, S.; Bhagavatula, C.; Zellers, R.; Peters, M. E.; Sabharwal, A.; and Choi, Y. 2020 · 2002
Earlier work this paper cites.
Automatically constructing a corpus of sentential paraphrases
Dolan, W. B.; and Brockett, C. 2005 · 2005
Earlier work this paper cites.
Pruksachatkun, Y.; Phang, J.; Liu, H.; Htut, P. M.; Zhang, X.; Pang, R. Y.; Vania, C.; Kann, K.; and Bowman, S. R. 2020 · 2005
Earlier work this paper cites.
Exploring and predicting transferability across nlp tasks
Vu, T.; Wang, T.; Munkhdalai, T.; Sordoni, A.; Trischler, A.; Mattarella-Micke, A.; Maji, S.; and Iyyer, M. 2020 · 2005
Earlier work this paper cites.
The second PASCAL recognising textual entailment challenge
Bar Haim, R.; Dagan, I.; Dolan, B.; Ferro, L.; Giampiccolo, D.; Magnini, B.; and Szpektor, I. 2006 · 2006
Earlier work this paper cites.
The PASCAL recognising textual entailment challenge
Dagan, I.; Glickman, O.; and Magnini, B. 2006 · 2006
Earlier work this paper cites.
Proceedings of the Fourth International Workshop on Semantic Evaluations (SemEval-2007)
Agirre, E.; M‘arquez, L.; and Wicentowski, R., eds. 2007 · 2007
Earlier work this paper cites.
The third PASCAL recognizing textual entailment challenge
Giampiccolo, D.; Magnini, B.; Dagan, I.; and Dolan, B. 2007 · 2007
Earlier work this paper cites.
The Fifth PASCAL Recognizing Textual Entailment Challenge
Bentivogli, L.; Dagan, I.; Dang, H. T.; Giampiccolo, D.; and Magnini, B. 2009 · 2009
Earlier work this paper cites.
The Winograd schema challenge
Levesque, H. J.; Davis, E.; and Morgenstern, L. 2011 · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
Pedregosa, F.; Varoquaux, G.; Gramfort, A.; Michel, V.; Thirion, B.; Grisel, O.; Blondel, M.; Prettenhofer, P.; Weiss, R.; Dubourg, V.; et al. 2011 · 2011
Cited alongside, same era.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Roemmele, M.; Bejan, C. A.; and Gordon, A. S. 2011 · 2011
Cited alongside, same era.
GNU Parallel - The Command-Line Power Tool
Tange, O. 2011 · 2011
Cited alongside, same era.
Reporting bias and knowledge acquisition
Gordon, J.; and Van Durme, B. 2013 · 2013
Cited alongside, same era.
Distributed Representations of Words and Phrases and their Compositionality
Mikolov, T.; Sutskever, I.; Chen, K.; Corrado, G. S.; and Dean, J. 2013 · 2013
Cited alongside, same era.
Recursive deep models for semantic compositionality over a sentiment treebank
Socher, R.; Perelygin, A.; Wu, J.; Chuang, J.; Manning, C. D.; Ng, A.; and Potts, C. 2013 · 2013
A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference
Williams, A.; Nangia, N.; and Bowman, S. R. 2018 · 2018
Later among the works it cites.
SWAG: A Large-Scale Adversarial Dataset for Grounded Commonsense Inference
Zellers, R.; Bisk, Y.; Schwartz, R.; and Choi, Y. 2018 · 2018
Later among the works it cites.
ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension
Zhang, S.; Liu, X.; Liu, J.; Gao, J.; Duh, K.; and Durme, B. V. 2018 · 2018
Later among the works it cites.
COMET: Commonsense Transformers for Automatic Knowledge Graph Construction
Bosselut, A.; Rashkin, H.; Sap, M.; Malaviya, C.; Çelikyilmaz, A.; and Choi, Y. 2019 · 2019
Later among the works it cites.
BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Clark, C.; Lee, K.; Chang, M.-W.; Kwiatkowski, T.; Collins, M.; and Toutanova, K. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
TensorFlow: A system for large-scale machine learning
Abadi, M.; Barham, P.; Chen, J.; Chen, Z.; Davis, A.; Dean, J.; Devin, M.; Ghemawat, S.; Irving, G.; Isard, M.; Kudlur, M.; Levenberg, J.; Monga, R.; Moore, S.; Murray, D. G.; Steiner, B.; Tucker, P.; Vasudevan, V.; Warden, P.; Wicke, M.; Yu, Y.; and Zheng, X. 2016 · 2016
Cited alongside, same era.
A Corpus and Cloze Evaluation for Deeper Understanding of Commonsense Stories
Mostafazadeh, N.; Chambers, N.; He, X.; Parikh, D.; Batra, D.; Vanderwende, L.; Kohli, P.; and Allen, J. 2016 · 2016
Cited alongside, same era.
SQuAD: 100,000+ Questions for Machine Comprehension of Text
Rajpurkar, P.; Zhang, J.; Lopyrev, K.; and Liang, P. 2016 · 2016
Cited alongside, same era.
Annotation Artifacts in Natural Language Inference Data
Gururangan, S.; Swayamdipta, S.; Levy, O.; Schwartz, R.; Bowman, S.; and Smith, N. A. 2018 · 2017
Cited alongside, same era.
Deep Learning Scaling is Predictable, Empirically
Hestness, J.; Narang, S.; Ardalani, N.; Diamos, G. F.; Jun, H.; Kianinejad, H.; Patwary, M. M. A.; Yang, Y.; and Zhou, Y. 2017 · 2017
Cited alongside, same era.
ConceptNet 5.5: An Open Multilingual Graph of General Knowledge
Speer, R.; Chin, J.; and Havasi, C. 2017 · 2017
Cited alongside, same era.
De Marneffe, M.-C.; Simons, M.; and Tonhauser, J. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Later among the works it cites.
Show Your Work: Improved Reporting of Experimental Results
Dodge, J.; Gururangan, S.; Card, D.; Schwartz, R.; and Smith, N. A. 2019 · 2019
Later among the works it cites.
Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning
Huang, L.; Le Bras, R.; Bhagavatula, C.; and Choi, Y. 2019 · 2019
Later among the works it cites.
Towards Generalizable Neuro-Symbolic Systems for Commonsense Question Answering
Ma, K.; Francis, J.; Lu, Q.; Nyberg, E.; and Oltramari, A. 2019 · 2019
Later among the works it cites.
WiC: The Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations
Pilehvar, M. T.; and Camacho-Collados, J. 2019 · 2019
Later among the works it cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2019 · 2019
Later among the works it cites.
Social IQA: Commonsense Reasoning about Social Interactions
Sap, M.; Rashkin, H.; Chen, D.; Le Bras, R.; and Choi, Y. 2019b · 2019
Later among the works it cites.
The Bitter Lesson
Sutton, R. S. 2019 · 2019
Later among the works it cites.
CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge
Talmor, A.; Herzig, J.; Lourie, N.; and Berant, J. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z.; Dai, Z.; Yang, Y.; Carbonell, J.; Salakhutdinov, R. R.; and Le, Q. V. 2019 · 2019
Later among the works it cites.
HellaSwag: Can a Machine Really Finish Your Sentence?
Zellers, R.; Holtzman, A.; Bisk, Y.; Farhadi, A.; and Choi, Y. 2019 · 2019
Later among the works it cites.
Abductive commonsense reasoning
Bhagavatula, C.; Le Bras, R.; Malaviya, C.; Sakaguchi, K.; Holtzman, A.; Rashkin, H.; Downey, D.; Yih, S. W.-t.; and Choi, Y. 2020 · 2020
Later among the works it cites.
PIQA: Reasoning about Physical Commonsense in Natural Language
Bisk, Y.; Zellers, R.; Le Bras, R.; Gao, J.; and Choi, Y. 2020 · 2020
Later among the works it cites.
UnifiedQA: Crossing Format Boundaries With a Single QA System
Khashabi, D.; Min, S.; Khot, T.; Sabhwaral, A.; Tafjord, O.; Clark, P.; and Hajishirzi, H. 2020 · 2020
Later among the works it cites.
A Constructive Prediction of the Generalization Error Across Scales
Rosenfeld, J. S.; Rosenfeld, A.; Belinkov, Y.; and Shavit, N. 2020 · 2020
Later among the works it cites.
WINOGRANDE: An Adversarial Winograd Schema Challenge at Scale
Sakaguchi, K.; Le Bras, R.; Bhagavatula, C.; and Choi, Y. 2020 · 2020
Later among the works it cites.
L2R²: Leveraging Ranking for Abductive Reasoning
Zhu, Y.; Pang, L.; Lan, Y.; and Cheng, X. 2020 · 2020
Later among the works it cites.