Fetching the paper…
Reading the bibliography…
Reasoning, a crucial ability for complex problem-solving, plays a pivotal role in various real-world settings such as negotiation, medical diagnosis, and criminal investigation.
Brown T, Mann B, Ryder N, et al. (2020) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
1906
Earlier work this paper cites.
1910
Earlier work this paper cites.
Fayek HM, Johnson J (2019) Temporal reasoning via audio question answering. 1911.09655
1911
Earlier work this paper cites.
Raven JC, Court J (1938) Raven’s progressive matrices. Western Psychological Services Los Angeles, CA
1938
Earlier work this paper cites.
Reiter R (1975) Formal reasoning and language understanding system. In: Theoretical Issues in Natural Language Processing
1975
Earlier work this paper cites.
Berzonsky MD (1978) Formal reasoning in adolescence: An alternative view. Adolescence 13(50):279
1978
Earlier work this paper cites.
Kahneman D, Miller DT (1986) Norm theory: Comparing reality to its alternatives. Psychological review 93(2):136
1986
Earlier work this paper cites.
Pollock JL (1987) Defeasible reasoning. Cognitive science 11(4):481–518
1987
Earlier work this paper cites.
Collins A, Michalski R (1989) The logic of plausible reasoning: A core theory. cognitive science 13(1):1–49
1989
Earlier work this paper cites.
Salmon W, Salmon M, Kitcher P (1989) Scientific explanation. Minneapolis
1989
Earlier work this paper cites.
Hemphill CT, Godfrey JJ, Doddington GR (1990) The ATIS spoken language systems pilot corpus. In: Speech and Natural Language: Proceedings of a Workshop Held at Hidden Valley, Pennsylvania, June 24-27,1990, URL https://aclanthology.org/H90-1021
1990
Earlier work this paper cites.
Hinton GE (1990) Connectionist learning procedures. In: Machine learning. Elsevier, p 555–610
1990
Earlier work this paper cites.
Jacobs RA, Jordan MI, Nowlan SJ, et al. (1991) Adaptive mixtures of local experts. Neural computation 3(1):79–87
1991
Earlier work this paper cites.
Pollock JL (1991) A theory of defeasible reasoning. International Journal of Intelligent Systems 6(1):33–54
1991
Earlier work this paper cites.
Paulson LC (1994) Isabelle: A generic theorem prover. Springer
1994
Earlier work this paper cites.
Carpenter TP, Fennema E, Franke ML (1996) Cognitively guided instruction: A knowledge base for reform in primary mathematics instruction. The elementary school journal 97(1):3–20
1996
Earlier work this paper cites.
Fennema E, Carpenter TP, Franke ML, et al. (1996) A longitudinal study of learning to use children’s thinking in mathematics instruction. Journal for research in mathematics education 27(4):403–434
1996
Earlier work this paper cites.
Oberlander J, Cox R, Stenning K (1996) Proof styles in multimodal reasoning. In: Logic, Language and Computation. CSLI Publications, p 403–414
1996
Earlier work this paper cites.
Barras B, Boutin S, Cornes C, et al. (1997) The coq proof assistant reference manual: Version 6.1. PhD thesis, Inria
1997
Earlier work this paper cites.
Cohen GH (1997) Align: a program to superimpose protein coordinates, accounting for insertions and deletions. Journal of applied crystallography 30(6):1160–1161
1997
Earlier work this paper cites.
Byrne RM, Tasso A (1999) Deductive reasoning with factual, possible, and counterfactual conditionals. Memory & cognition 27:726–740
1999
Earlier work this paper cites.
Flach PA, Kakas AC (2000) Abductive and Inductive Reasoning: Background and Issues, Springer Netherlands, Dordrecht, pp 1–27. 10.1007/978-94-017-0606-3_1 , URL https://doi.org/10.1007/978-94-017-0606-3_1
2000
Earlier work this paper cites.
Papineni K, Roukos S, Ward T, et al. (2002) Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics, pp 311–318
2002
Earlier work this paper cites.
Strehl A, Ghosh J (2002) Cluster ensembles—a knowledge reuse framework for combining multiple partitions. Journal of machine learning research 3(Dec):583–617
2002
Earlier work this paper cites.
Sowa JF (2003) Laws, facts, and contexts: Foundations for multimodal reasoning. In: Knowledge Contributors. Springer, p 145–184
2003
Earlier work this paper cites.
Berghofer S, Strecker M (2004) Extracting a formally verified, fully executable compiler from a proof assistant. Electronic Notes in Theoretical Computer Science 82(2):377–394
2004
Earlier work this paper cites.
Evans JSB, Thompson VA (2004) Informal reasoning: Theory and method. Canadian Journal of Experimental Psychology/Revue canadienne de psychologie expérimentale 58(2):69
2004
Earlier work this paper cites.
Floyd J (2004) Wittgenstein on philosophy of logic and mathematics. Graduate Faculty Philosophy Journal 25(2):227–287
2004
Earlier work this paper cites.
Redbooks I (2004) Practical Guide to the IBM Autonomic Computing Toolkit. IBM
2004
Earlier work this paper cites.
Koons R (2005) Defeasible reasoning. arXiv
2005
Earlier work this paper cites.
Li L, Szygenda SA, Thornton MA (2005) Combining simulation and formal verification for integrated circuit design validation. In: Proceedings of the 9th World Multi-Conference on Systemics, Cybernetics and Informatics (WMSCI), pp 92–97
2005
Earlier work this paper cites.
Acay DL, Pasquier P, Sonenberg L (2007) Extrospection: Agents reasoning about the environment. 3rd IET International Conference on Intelligent Environments
2007
Earlier work this paper cites.
Byrne RM (2007) The rational imagination: How people create alternatives to reality. MIT press
2007
Earlier work this paper cites.
Garcez AS, Lamb LC, Gabbay DM (2008) Neural-symbolic cognitive reasoning. Springer Science & Business Media
2008
Earlier work this paper cites.
Nelson B, Barreno M, Chi FJ, et al. (2008) Exploiting machine learning to subvert your spam filter. LEET 8(1-9):16–17
2008
Earlier work this paper cites.
Deng J, Dong W, Socher R, et al. (2009) Imagenet: A large-scale hierarchical image database. In: 2009 IEEE conference on computer vision and pattern recognition, Ieee, pp 248–255
2009
Earlier work this paper cites.
Pollock JL (2009) A recursive semantics for defeasible reasoning. Argumentation in artificial intelligence pp 173–197
2009
Earlier work this paper cites.
Woodcock J, Larsen PG, Bicarregui J, et al. (2009) Formal methods: Practice and experience. ACM computing surveys (CSUR) 41(4):1–36
2009
Earlier work this paper cites.
De Raedt L, Kersting K (2010) Statistical relational learning. Encyclopedia of Machine Learning
2010
Earlier work this paper cites.
Harrison J (2010) Formal methods at intel—an overview. In: Second NASA Formal Methods Symposium, pp 179–195
2010
Earlier work this paper cites.
Lin Z, Wu YF, Peri S, et al. (2020c) Improving generative imagination in object-centric world models. 2010.02054
2010
Earlier work this paper cites.
Paulson LC (2010) Three years of experience with sledgehammer, a practical link between automatic and interactive theorem provers. In: Schmidt RA, Schulz S, Konev B (eds) Proceedings of the 2nd Workshop on Practical Aspects of Automated Reasoning, PAAR-2010, Edinburgh, Scotland, UK, July 14, 2010, EPiC Series in Computing, vol 9. EasyChair, pp 1–10, 10.29007/tnfd , URL https://doi.org/10.29007/tnfd
2010
Earlier work this paper cites.
2010
Earlier work this paper cites.
Brandt A, McClure R (2011) Sound reasoning
2011
Earlier work this paper cites.
Ordonez V, Kulkarni G, Berg T (2011) Im2text: Describing images using 1 million captioned photographs. Advances in neural information processing systems 24
2011
Earlier work this paper cites.
Brewka G (2012) Default Reasoning, Springer US, Boston, MA, pp 915–917. 10.1007/978-1-4419-1428-6_634 , URL https://doi.org/10.1007/978-1-4419-1428-6_634
2012
Earlier work this paper cites.
Leake DB (2012) Introspective Learning and Reasoning, Springer US, Boston, MA, pp 1638–1640. 10.1007/978-1-4419-1428-6_1802 , URL https://doi.org/10.1007/978-1-4419-1428-6_1802
2012
Earlier work this paper cites.
Morris BJ, Croker S, Masnick AM, et al. (2012) The emergence of scientific reasoning. In: Kloos H, Morris BJ, Amaral JL (eds) Current Topics in Children’s Learning and Cognition. IntechOpen, Rijeka, chap 4, 10.5772/53885 , URL https://doi.org/10.5772/53885
2012
Earlier work this paper cites.
Berant J, Chou A, Frostig R, et al. (2013) Semantic parsing on freebase from question-answer pairs. In: Proceedings of the 2013 conference on empirical methods in natural language processing, pp 1533–1544
2013
Earlier work this paper cites.
Bottou L, Peters J, Quiñonero-Candela J, et al. (2013) Counterfactual reasoning and learning systems: The example of computational advertising. Journal of Machine Learning Research 14(11)
2013
Earlier work this paper cites.
Cai Q, Yates A (2013) Large-scale semantic parsing via schema matching and lexicon extension. In: Proceedings of the 51st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Sofia, Bulgaria, pp 423–433, URL https://aclanthology.org/P13-1042
2013
Earlier work this paper cites.
Waldmann MR, Hagmayer Y (2013) Causal reasoning. arXiv
2013
Earlier work this paper cites.
Alibali MW, Boncoddo R, Hostetter AB (2014) Gesture in reasoning: An embodied perspective. In: The Routledge handbook of embodied cognition. Routledge, p 150–159
2014
Earlier work this paper cites.
Hosseini MJ, Hajishirzi H, Etzioni O, et al. (2014) Learning to solve arithmetic word problems with verb categorization. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP). Association for Computational Linguistics, Doha, Qatar, pp 523–533, 10.3115/v1/D14-1058 , URL https://aclanthology.org/D14-1058
2014
Earlier work this paper cites.
Kushman N, Artzi Y, Zettlemoyer L, et al. (2014) Learning to automatically solve algebra word problems. In: Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Baltimore, Maryland, pp 271–281, 10.3115/v1/P14-1026 , URL https://aclanthology.org/P14-1026
2014
Earlier work this paper cites.
Lahiri S (2014) Complexity of Word Collocation Networks: A Preliminary Structural Analysis. In: Proceedings of the Student Research Workshop at the 14th Conference of the European Chapter of the Association for Computational Linguistics. Association for Computational Linguistics, Gothenburg, Sweden, pp 96–105, URL http://www.aclweb.org/anthology/E14-3011
2014
Earlier work this paper cites.
Al-Ajlan A (2015) The comparison between forward and backward chaining. International Journal of Machine Learning and Computing 5(2):106
2015
Earlier work this paper cites.
Garcez Ad, Besold TR, De Raedt L, et al. (2015) Neural-symbolic learning and reasoning: contributions and challenges. In: 2015 AAAI Spring Symposium Series
2015
Earlier work this paper cites.
Koncel-Kedziorski R, Hajishirzi H, Sabharwal A, et al. (2015) Parsing algebraic word problems into equations. Transactions of the Association for Computational Linguistics 3:585–597. 10.1162/tacl_a_00160 , URL https://aclanthology.org/Q15-1042
2015
Earlier work this paper cites.
de Moura L, Kong S, Avigad J, et al. (2015) The lean theorem prover (system description). In: Automated Deduction-CADE-25: 25th International Conference on Automated Deduction, Berlin, Germany, August 1-7, 2015, Proceedings 25, Springer, pp 378–388
2015
Earlier work this paper cites.
Pasupat P, Liang P (2015) Compositional semantic parsing on semi-structured tables. In: Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers). Association for Computational Linguistics, Beijing, China, pp 1470–1480, 10.3115/v1/P15-1142 , URL https://aclanthology.org/P15-1142
2015
Earlier work this paper cites.
Schmidhuber J (2015) Deep learning in neural networks: An overview. Neural networks 61:85–117
2015
Earlier work this paper cites.
Seo M, Hajishirzi H, Farhadi A, et al. (2015) Solving geometry problems: Combining text and diagram interpretation. In: Màrquez L, Callison-Burch C, Su J (eds) Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Lisbon, Portugal, pp 1466–1476, 10.18653/v1/D15-1171 , URL https://aclanthology.org/D15-1171
2015
Earlier work this paper cites.
Shi S, Wang Y, Lin CY, et al. (2015) Automatically solving number word problems by semantic parsing and reasoning. In: Conference on Empirical Methods in Natural Language Processing
2015
Earlier work this paper cites.
Upadhyay S, Chang MW (2015) Draw: A challenging and diverse algebra word problem set
2015
Earlier work this paper cites.
Vedantam R, Zitnick CL, Parikh D (2015) Cider: Consensus-based image description evaluation. 1411.5726
2015
Earlier work this paper cites.
Zhou L, Dai S, Chen L (2015) Learn to solve algebra word problems using quadratic programming. In: Conference on Empirical Methods in Natural Language Processing
2015
Earlier work this paper cites.
do Nascimento NM, de Lucena CJP (2017) Fiot: An agent-based framework for self-adaptive and self-organizing applications based on the internet of things. Information Sciences 378:161–176. https://doi.org/10.1016/j.ins.2016.10.031 , URL https://www.sciencedirect.com/science/article/pii/S0020025516313664
2016
Earlier work this paper cites.
Halpern JY (2016) Actual causality. MiT Press
2016
Earlier work this paper cites.
Huang D, Shi S, Lin CY, et al. (2016) How well do computers solve math word problems? large-scale dataset construction and evaluation. In: Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp 887–896
2016
Earlier work this paper cites.
Jaderberg M, Simonyan K, Zisserman A, et al. (2016) Spatial transformer networks. 1506.02025
2016
Earlier work this paper cites.
Koncel-Kedziorski R, Roy S, Amini A, et al. (2016) Mawps: A math word problem repository. In: North American Chapter of the Association for Computational Linguistics
2016
Earlier work this paper cites.
Miller A, Fisch A, Dodge J, et al. (2016) Key-value memory networks for directly reading documents. In: Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Austin, Texas, pp 1400–1409, 10.18653/v1/D16-1147 , URL https://aclanthology.org/D16-1147
2016
Earlier work this paper cites.
Mooij JM, Peters J, Janzing D, et al. (2016) Distinguishing cause from effect using observational data: methods and benchmarks. The Journal of Machine Learning Research 17(1):1103–1204
2016
Earlier work this paper cites.
Rajpurkar P, Zhang J, Lopyrev K, et al. (2016) SQuAD: 100,000+ questions for machine comprehension of text. In: Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Austin, Texas, pp 2383–2392, 10.18653/v1/D16-1264 , URL https://aclanthology.org/D16-1264
2016
Earlier work this paper cites.
Silver D, Huang A, Maddison CJ, et al. (2016) Mastering the game of go with deep neural networks and tree search. nature 529(7587):484–489
2016
Earlier work this paper cites.
Teig N, Scherer R (2016) Bringing formal and informal reasoning together—a new era of assessment? Frontiers in psychology 7:1097
2016
Earlier work this paper cites.
Thomee B, Shamma DA, Friedland G, et al. (2016) Yfcc100m: The new data in multimedia research. Communications of the ACM 59(2):64–73
2016
Earlier work this paper cites.
Alvin C, Gulwani S, Majumdar R, et al. (2017) Synthesis of solutions for shaded area geometry problems. In: The Thirtieth International Flairs Conference
2017
Earlier work this paper cites.
Arandjelovic R, Zisserman A (2017) Look, listen and learn. In: Proceedings of the IEEE international conference on computer vision, pp 609–617
2017
Earlier work this paper cites.
Bengio Y (2017) The consciousness prior. arXiv preprint arXiv:170908568
2017
Earlier work this paper cites.
Daniel K (2017) Thinking, fast and slow
2017
Earlier work this paper cites.
Gemmeke JF, Ellis DP, Freedman D, et al. (2017) Audio set: An ontology and human-labeled dataset for audio events. In: 2017 IEEE international conference on acoustics, speech and signal processing (ICASSP), IEEE, pp 776–780
2017
Earlier work this paper cites.
Johnson J, Hariharan B, Van Der Maaten L, et al. (2017) Clevr: A diagnostic dataset for compositional language and elementary visual reasoning. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 2901–2910
2017
Earlier work this paper cites.
Joshi M, Choi E, Weld D, et al. (2017) TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Vancouver, Canada, pp 1601–1611, 10.18653/v1/P17-1147 , URL https://aclanthology.org/P17-1147
2017
Earlier work this paper cites.
Li J, Seltzer ML, Wang X, et al. (2017) Large-scale domain adaptation via teacher-student learning. In: INTERSPEECH, pp 2386–2390
2017
Earlier work this paper cites.
Ling W, Yogatama D, Dyer C, et al. (2017) Program induction by rationale generation: Learning to solve and explain algebraic word problems. In: Barzilay R, Kan MY (eds) Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Vancouver, Canada, pp 158–167, 10.18653/v1/P17-1015 , URL https://aclanthology.org/P17-1015
2017
Earlier work this paper cites.
Peters J, Janzing D, Schölkopf B (2017) Elements of causal inference: foundations and learning algorithms. The MIT Press
2017
Earlier work this paper cites.
Sachan M, Xing E (2017) Learning to solve geometry problems from natural language demonstrations in textbooks. In: Proceedings of the 6th Joint Conference on Lexical and Computational Semantics (*SEM 2017). Association for Computational Linguistics, Vancouver, Canada, pp 251–261, 10.18653/v1/S17-1029 , URL https://aclanthology.org/S17-1029
2017
Earlier work this paper cites.
Sachan M, Dubey K, Xing E (2017) From textbooks to knowledge: A case study in harvesting axiomatic knowledge from textbooks to solve geometry problems. In: Palmer M, Hwa R, Riedel S (eds) Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Copenhagen, Denmark, pp 773–784, 10.18653/v1/D17-1081 , URL https://aclanthology.org/D17-1081
2017
Earlier work this paper cites.
Schulman J, Wolski F, Dhariwal P, et al. (2017) Proximal policy optimization algorithms. arXiv preprint arXiv:170706347
2017
Earlier work this paper cites.
Shazeer N, Mirhoseini A, Maziarz K, et al. (2017) Outrageously large neural networks: The sparsely-gated mixture-of-experts layer. arXiv preprint arXiv:170106538
2017
Earlier work this paper cites.
Speer R, Chin J, Havasi C (2017) Conceptnet 5.5: An open multilingual graph of general knowledge. In: Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence. AAAI Press, AAAI’17, p 4444–4451
2017
Earlier work this paper cites.
Sun C, Shrivastava A, Singh S, et al. (2017) Revisiting unreasonable effectiveness of data in deep learning era. In: Proceedings of the IEEE international conference on computer vision, pp 843–852
2017
Earlier work this paper cites.
Upadhyay S, Chang MW (2017) Annotating derivations: A new evaluation strategy and dataset for algebra word problems. In: Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics: Volume 1, Long Papers. Association for Computational Linguistics, Valencia, Spain, pp 494–504, URL https://aclanthology.org/E17-1047
2017
Earlier work this paper cites.
Van Den Oord A, Vinyals O, et al. (2017) Neural discrete representation learning. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Vaswani A, Shazeer N, Parmar N, et al. (2017) Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Wang Y, Liu X, Shi S (2017) Deep neural solver for math word problems. In: Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Copenhagen, Denmark, pp 845–854, 10.18653/v1/D17-1088 , URL https://aclanthology.org/D17-1088
2017
Earlier work this paper cites.
Zhang Y, Dai H, Kozareva Z, et al. (2017) Variational reasoning for question answering with knowledge graph. In: AAAI Conference on Artificial Intelligence
2017
Earlier work this paper cites.
Dong L, Xu S, Xu B (2018) Speech-transformer: a no-recurrence sequence-to-sequence model for speech recognition. In: ICASSP, pp 5884–5888
2018
Earlier work this paper cites.
Oord Avd, Li Y, Vinyals O (2018) Representation learning with contrastive predictive coding. arXiv preprint arXiv:180703748
2018
Earlier work this paper cites.
Puig X, Ra K, Boben M, et al. (2018) Virtualhome: Simulating household activities via programs. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 8494–8502
2018
Earlier work this paper cites.
Radford A, Narasimhan K, Salimans T, et al. (2018) Improving language understanding by generative pre-training. arXiv
2018
Earlier work this paper cites.
Sharma P, Ding N, Goodman S, et al. (2018) Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp 2556–2565
2018
Earlier work this paper cites.
Silver D, Hubert T, Schrittwieser J, et al. (2018) A general reinforcement learning algorithm that masters chess, shogi, and go through self-play. Science 362(6419):1140–1144
2018
Earlier work this paper cites.
Song X, Shi Y, Chen X, et al. (2018) Explore multi-step reasoning in video question answering. In: Proceedings of the 26th ACM international conference on Multimedia, pp 239–247
2018
Earlier work this paper cites.
Talmor A, Berant J (2018) The web as a knowledge-base for answering complex questions. In: Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers). Association for Computational Linguistics, New Orleans, Louisiana, pp 641–651, 10.18653/v1/N18-1059 , URL https://aclanthology.org/N18-1059
2018
Earlier work this paper cites.
Vilares D, Peng H, Satapathy R, et al. (2018) Babelsenticnet: a commonsense reasoning framework for multilingual sentiment analysis. In: 2018 IEEE symposium series on computational intelligence (SSCI), IEEE, pp 1292–1298
2018
Earlier work this paper cites.
Xia F, R. Zamir A, He ZY, et al. (2018) Gibson Env: real-world perception for embodied agents. In: Computer Vision and Pattern Recognition (CVPR), 2018 IEEE Conference on, IEEE
2018
Earlier work this paper cites.
Yang Z, Qi P, Zhang S, et al. (2018) HotpotQA: A dataset for diverse, explainable multi-hop question answering. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Brussels, Belgium, pp 2369–2380, 10.18653/v1/D18-1259 , URL https://aclanthology.org/D18-1259
2018
Earlier work this paper cites.
Yu T, Zhang R, Yang K, et al. (2018) Spider: A large-scale human-labeled dataset for complex and cross-domain semantic parsing and text-to-SQL task. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Brussels, Belgium, pp 3911–3921, 10.18653/v1/D18-1425 , URL https://aclanthology.org/D18-1425
2018
Earlier work this paper cites.
Zellers R, Bisk Y, Schwartz R, et al. (2018) Swag: A large-scale adversarial dataset for grounded commonsense inference. arXiv preprint arXiv:180805326
2018
Earlier work this paper cites.
Zhong V, Xiong C, Socher R (2018) Seq2SQL: Generating structured queries from natural language using reinforcement learning. URL https://openreview.net/forum?id=Syx6bz-Ab
2018
Earlier work this paper cites.
Amini A, Gabriel S, Lin P, et al. (2019) Mathqa: Towards interpretable math word problem solving with operation-based formalisms. arXiv preprint arXiv:190513319
2019
Earlier work this paper cites.
Araci D (2019) Finbert: Financial sentiment analysis with pre-trained language models. arXiv preprint arXiv:190810063
2019
Earlier work this paper cites.
Bakhtin A, van der Maaten L, Johnson J, et al. (2019) Phyre: A new benchmark for physical reasoning. Advances in Neural Information Processing Systems 32
2019
Earlier work this paper cites.
Bansal K, Loos S, Rabe M, et al. (2019) HOList: An environment for machine learning of higher order logic theorem proving. In: Chaudhuri K, Salakhutdinov R (eds) Proceedings of the 36th International Conference on Machine Learning, Proceedings of Machine Learning Research, vol 97. PMLR, pp 454–463, URL https://proceedings.mlr.press/v97/bansal19a.html
2019
Earlier work this paper cites.
Bhagavatula C, Bras RL, Malaviya C, et al. (2019) Abductive commonsense reasoning. arXiv preprint arXiv:190805739
2019
Earlier work this paper cites.
Burgess CP, Matthey L, Watters N, et al. (2019) Monet: Unsupervised scene decomposition and representation. arXiv preprint arXiv:190111390
2019
Earlier work this paper cites.
Chen J, Lin St, Durrett G (2019) Multi-hop question answering via reasoning chains. arXiv preprint arXiv:191002610
2019
Earlier work this paper cites.
Chung YA, Hsu WN, Tang H, et al. (2019) An unsupervised autoregressive model for speech representation learning. In: INTERSPEECH, pp 146–150
2019
Earlier work this paper cites.
Devlin J, Chang MW, Lee K, et al. (2019) BERT: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, pp 4171–4186, 10.18653/v1/N19-1423 , URL https://aclanthology.org/N19-1423
2019
Earlier work this paper cites.
Dua D, Wang Y, Dasigi P, et al. (2019) DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, pp 2368–2378, 10.18653/v1/N19-1246 , URL https://aclanthology.org/N19-1246
2019
Earlier work this paper cites.
Furbach U, Hölldobler S, Ragni M, et al. (2019) Cognitive reasoning: A personal view. KI-Künstliche Intelligenz 33:209–217
2019
Earlier work this paper cites.
Gupta N, Lin K, Roth D, et al. (2019) Neural module networks for reasoning over text. arXiv preprint arXiv:191204971
2019
Earlier work this paper cites.
Houlsby N, Giurgiu A, Jastrzebski S, et al. (2019) Parameter-efficient transfer learning for nlp. In: International Conference on Machine Learning, PMLR, pp 2790–2799
2019
Earlier work this paper cites.
Huang L, Bras RL, Bhagavatula C, et al. (2019) Cosmos qa: Machine reading comprehension with contextual commonsense reasoning. arXiv preprint arXiv:190900277
2019
Earlier work this paper cites.
Hudson DA, Manning CD (2019) Gqa: A new dataset for real-world visual reasoning and compositional question answering. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp 6700–6709
2019
Earlier work this paper cites.
Kenton JDMWC, Toutanova LK (2019) Bert: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of naacL-HLT, p 2
2019
Earlier work this paper cites.
Kim J, Misu T, Chen YT, et al. (2019) Grounding human-to-vehicle advice for self-driving vehicles. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition
2019
Earlier work this paper cites.
Kwiatkowski T, Palomaki J, Redfield O, et al. (2019) Natural questions: A benchmark for question answering research. Transactions of the Association for Computational Linguistics 7:452–466. 10.1162/tacl_a_00276 , URL https://aclanthology.org/Q19-1026
2019
Earlier work this paper cites.
Liu Y, Ott M, Goyal N, et al. (2019) Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:190711692
2019
Earlier work this paper cites.
Manolis Savva*, Abhishek Kadian*, Oleksandr Maksymets*, et al. (2019) Habitat: A Platform for Embodied AI Research. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)
2019
Earlier work this paper cites.
Mao J, Gan C, Kohli P, et al. (2019) The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision. In: International Conference on Learning Representations, URL https://openreview.net/forum?id=rJgMlhRctm
2019
Earlier work this paper cites.
Marino K, Rastegari M, Farhadi A, et al. (2019) Ok-vqa: A visual question answering benchmark requiring external knowledge. In: Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Earlier work this paper cites.
Megill N, Wheeler DA (2019) A computer language for mathematical proofs. arXiv
2019
Earlier work this paper cites.
Radford A, Wu J, Child R, et al. (2019) Language models are unsupervised multitask learners. OpenAI blog 1(8):9
2019
Earlier work this paper cites.
Reimers N, Gurevych I (2019) Sentence-bert: Sentence embeddings using siamese bert-networks. arXiv preprint arXiv:190810084
2019
Earlier work this paper cites.
Sap M, Rashkin H, Chen D, et al. (2019) Social IQa: Commonsense reasoning about social interactions. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). Association for Computational Linguistics, Hong Kong, China, pp 4463–4473, 10.18653/v1/D19-1454 , URL https://aclanthology.org/D19-1454
2019
Earlier work this paper cites.
Sinha K, Sodhani S, Dong J, et al. (2019) Clutrr: A diagnostic benchmark for inductive reasoning from text. arXiv preprint arXiv:190806177
2019
Earlier work this paper cites.
Talmor A, Herzig J, Lourie N, et al. (2019) CommonsenseQA: A question answering challenge targeting commonsense knowledge. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, pp 4149–4158, 10.18653/v1/N19-1421 , URL https://aclanthology.org/N19-1421
2019
Earlier work this paper cites.
Tu R, Zhang K, Bertilson B, et al. (2019) Neuropathic pain diagnosis simulator for causal discovery algorithm evaluation. Advances in Neural Information Processing Systems 32
2019
Earlier work this paper cites.
Xie Z, Sun S (2019) A goal-driven tree-structured neural model for math word problems. In: Ijcai, pp 5299–5305
2019
Earlier work this paper cites.
Yang K, Deng J (2019) Learning to prove theorems via interacting with proof assistants. In: International Conference on Machine Learning, PMLR, pp 6984–6994
2019
Earlier work this paper cites.
Yi K, Gan C, Li Y, et al. (2019) Clevrer: Collision events for video representation and reasoning. arXiv preprint arXiv:191001442
2019
Earlier work this paper cites.
Zellers R, Holtzman A, Bisk Y, et al. (2019) Hellaswag: Can a machine really finish your sentence? arXiv preprint arXiv:190507830
2019
Earlier work this paper cites.
Ardila R, Branson M, Davis K, et al. (2020) Common Voice: A Massively-Multilingual Speech Corpus. In: Proceedings of the Twelfth Language Resources and Evaluation Conference, pp 4218–4222
2020
Earlier work this paper cites.
Asai A, Hashimoto K, Hajishirzi H, et al. (2020) Learning to retrieve reasoning paths over wikipedia graph for question answering. In: International Conference on Learning Representations, URL https://openreview.net/forum?id=SJgVHkrYDH
2020
Earlier work this paper cites.
Baevski A, Zhou Y, Mohamed A, et al. (2020) wav2vec 2.0: A framework for self-supervised learning of speech representations. Advances in neural information processing systems 33:12449–12460
2020
Earlier work this paper cites.
Berka P (2020) Sentiment analysis using rule-based and case-based reasoning. Journal of Intelligent Information Systems 55(1):51–66
2020
Earlier work this paper cites.
Bisk Y, Zellers R, Gao J, et al. (2020) Piqa: Reasoning about physical commonsense in natural language. In: Proceedings of the AAAI conference on artificial intelligence, pp 7432–7439
2020
Earlier work this paper cites.
Boratko M, Li XL, Das R, et al. (2020) Protoqa: A question answering dataset for prototypical common-sense reasoning. arXiv preprint arXiv:200500771
2020
Earlier work this paper cites.
Chen W, Zha H, Chen Z, et al. (2020a) HybridQA: A dataset of multi-hop question answering over tabular and textual data. In: Findings of the Association for Computational Linguistics: EMNLP 2020. Association for Computational Linguistics, Online, pp 1026–1036, 10.18653/v1/2020.findings-emnlp.91 , URL https://aclanthology.org/2020.findings-emnlp.91
2020
Earlier work this paper cites.
Clark P, Tafjord O, Richardson K (2020) Transformers as soft reasoners over language. arXiv preprint arXiv:200205867
2020
Earlier work this paper cites.
Clement CB, Drain D, Timcheck J, et al. (2020) Pymt5: multi-mode translation of natural language and python code with transformers. arXiv preprint arXiv:201003150
2020
Earlier work this paper cites.
Deitke M, Han W, Herrasti A, et al. (2020) Robothor: An open simulation-to-real embodied ai platform. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp 3164–3174
2020
Earlier work this paper cites.
Ding D, Hill F, Santoro A, et al. (2020) Object-based attention for spatio-temporal reasoning: Outperforming neuro-symbolic models with flexible distributed architectures. arXiv preprint arXiv:201208508 1
2020
Earlier work this paper cites.
Fang Y, Sun S, Gan Z, et al. (2020) Hierarchical graph network for multi-hop question answering. In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP). Association for Computational Linguistics, Online, pp 8823–8838, 10.18653/v1/2020.emnlp-main.710 , URL https://aclanthology.org/2020.emnlp-main.710
2020
Earlier work this paper cites.
Gao L, Biderman S, Black S, et al. (2020) The pile: An 800gb dataset of diverse text for language modeling. arXiv preprint arXiv:210100027
2020
Earlier work this paper cites.
Girdhar R, Ramanan D (2020) Cater: A diagnostic dataset for compositional actions and temporal reasoning. In: ICLR
2020
Earlier work this paper cites.
Gulati A, Qin J, Chiu CC, et al. (2020) Conformer: Convolution-augmented transformer for speech recognition. In: INTERSPEECH, pp 5036–5040
2020
Earlier work this paper cites.
Huang J, Xie S, Sun J, et al. (2021b) Learning a decision module by imitating driver’s control behaviors. In: Kober J, Ramos F, Tomlin C (eds) Proceedings of the 2020 Conference on Robot Learning, Proceedings of Machine Learning Research, vol 155. PMLR, pp 1–10, URL https://proceedings.mlr.press/v155/huang21a.html
2020
Earlier work this paper cites.
Jansen P (2020) Visually-grounded planning without vision: Language models infer detailed plans from high-level instructions. In: Findings of the Association for Computational Linguistics: EMNLP 2020. Association for Computational Linguistics, Online, pp 4412–4417, 10.18653/v1/2020.findings-emnlp.395 , URL https://aclanthology.org/2020.findings-emnlp.395
2020
Earlier work this paper cites.
Jiao R, Wang Z, Chu R, et al. (2020) An intuitive end-to-end human-uav interaction system for field exploration. Frontiers in Neurorobotics
2020
Earlier work this paper cites.
Kahn J, Rivière M, Zheng W, et al. (2020) Libri-light: A benchmark for asr with limited or no supervision. In: ICASSP, pp 7669–7673
2020
Earlier work this paper cites.
Khan W, Kamran M, Naqvi SR, et al. (2020) Formal verification of hardware components in critical systems. Wireless Communications and Mobile Computing 2020:1–15
2020
Earlier work this paper cites.
Khashabi D, Min S, Khot T, et al. (2020) Unifiedqa: Crossing format boundaries with a single qa system. arXiv preprint arXiv:200500700
2020
Earlier work this paper cites.
Kurita K, Michel P, Neubig G (2020) Weight poisoning attacks on pre-trained models. arXiv preprint arXiv:200406660
2020
Earlier work this paper cites.
Lepikhin D, Lee H, Xu Y, et al. (2020) Gshard: Scaling giant models with conditional computation and automatic sharding. arXiv preprint arXiv:200616668
2020
Earlier work this paper cites.
Lewis M, Liu Y, Goyal N, et al. (2020) BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. Association for Computational Linguistics, Online, pp 7871–7880, 10.18653/v1/2020.acl-main.703 , URL https://aclanthology.org/2020.acl-main.703
2020
Earlier work this paper cites.
Li X, Yin X, Li C, et al. (2020) Oscar: Object-semantics aligned pre-training for vision-language tasks. In: ECCV, pp 121–137
2020
Earlier work this paper cites.
Lin BY, Zhou W, Shen M, et al. (2020b) CommonGen: A constrained text generation challenge for generative commonsense reasoning. In: Findings of the Association for Computational Linguistics: EMNLP 2020. Association for Computational Linguistics, Online, pp 1823–1840, 10.18653/v1/2020.findings-emnlp.165 , URL https://aclanthology.org/2020.findings-emnlp.165
2020
Earlier work this paper cites.
Liu AT, Yang Sw, Chi PH, et al. (2020) Mockingjay: Unsupervised speech representation learning with deep bidirectional transformer encoders. In: ICASSP, pp 6419–6423
2020
Earlier work this paper cites.
Lo K, Wang LL, Neumann M, et al. (2020) S2ORC: The semantic scholar open research corpus. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. Association for Computational Linguistics, Online, pp 4969–4983, 10.18653/v1/2020.acl-main.447 , URL https://aclanthology.org/2020.acl-main.447
2020
Earlier work this paper cites.
Miao Sy, Liang CC, Su KY (2020) A diverse corpus for evaluating and developing english math word problem solvers. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pp 975–984
2020
Earlier work this paper cites.
Pan B, Sun J, Leung HYT, et al. (2020) Cross-view semantic segmentation for sensing surroundings. IEEE Robotics and Automation Letters 5(3):4867–4873. 10.1109/LRA.2020.3004325
2020
Earlier work this paper cites.
Pfeiffer J, Vulić I, Gurevych I, et al. (2020) Mad-x: An adapter-based framework for multi-task cross-lingual transfer. arXiv preprint arXiv:200500052
2020
Earlier work this paper cites.
Pilault J, Li R, Subramanian S, et al. (2020) On extractive and abstractive neural document summarization with transformer language models. In: Webber B, Cohn T, He Y, et al. (eds) Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP). Association for Computational Linguistics, Online, pp 9308–9319, 10.18653/v1/2020.emnlp-main.748 , URL https://aclanthology.org/2020.emnlp-main.748
2020
Earlier work this paper cites.
Polu S, Sutskever I (2020) Generative language modeling for automated theorem proving. arXiv preprint arXiv:200903393
2020
Earlier work this paper cites.
Pratap V, Xu Q, Sriram A, et al. (2020) MLS: A large-scale multilingual dataset for speech research. In: INTERSPEECH, pp 2757–2761
2020
Earlier work this paper cites.
Rajani NF, Zhang R, Tan YC, et al. (2020) ESPRIT: Explaining solutions to physical reasoning tasks. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. Association for Computational Linguistics, Online, pp 7906–7917, 10.18653/v1/2020.acl-main.706 , URL https://aclanthology.org/2020.acl-main.706
2020
Earlier work this paper cites.
Roller S, Dinan E, Goyal N, et al. (2020) Recipes for building an open-domain chatbot. arXiv preprint arXiv:200413637
2020
Earlier work this paper cites.
Rudinger R, Shwartz V, Hwang JD, et al. (2020) Thinking like a skeptic: Defeasible inference in natural language. In: Findings of the Association for Computational Linguistics: EMNLP 2020. Association for Computational Linguistics, Online, pp 4661–4675, 10.18653/v1/2020.findings-emnlp.418 , URL https://aclanthology.org/2020.findings-emnlp.418
2020
Earlier work this paper cites.
Shin T, Razeghi Y, Logan IV RL, et al. (2020) Autoprompt: Eliciting knowledge from language models with automatically generated prompts. arXiv preprint arXiv:201015980
2020
Earlier work this paper cites.
Sun J, Sun H, Han T, et al. (2021) Neuro-symbolic program search for autonomous driving decision module design. In: Kober J, Ramos F, Tomlin C (eds) Proceedings of the 2020 Conference on Robot Learning, Proceedings of Machine Learning Research, vol 155. PMLR, pp 21–30, URL https://proceedings.mlr.press/v155/sun21a.html
2020
Earlier work this paper cites.
Sun T, Shao Y, Qiu X, et al. (2020) CoLAKE: Contextualized language and knowledge embedding. In: Scott D, Bel N, Zong C (eds) Proceedings of the 28th International Conference on Computational Linguistics. International Committee on Computational Linguistics, Barcelona, Spain (Online), pp 3660–3670, 10.18653/v1/2020.coling-main.327 , URL https://aclanthology.org/2020.coling-main.327
2020
Earlier work this paper cites.
Svyatkovskiy A, Deng SK, Fu S, et al. (2020) Intellicode compose: Code generation using transformer. In: Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering, pp 1433–1443
2020
Cited alongside, same era.
Tu M, Huang K, Wang G, et al. (2020) Select, answer and explain: Interpretable multi-hop reading comprehension over multiple documents. In: Proceedings of the AAAI conference on artificial intelligence, pp 9073–9080
2020
Cited alongside, same era.
Wang Z, Wohlwend J, Lei T (2020) Structured pruning of large language models. In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP). Association for Computational Linguistics, 10.18653/v1/2020.emnlp-main.496 , URL https://doi.org/10.18653%2Fv1%2F2020.emnlp-main.496
2020
Cited alongside, same era.
Schuhmann C, Beaumont R, Vencu R, et al. (2022) Laion-5b: An open large-scale dataset for training next generation image-text models. Advances in Neural Information Processing Systems 35:25278–25294
2022
Later among the works it cites.
Shen S, Li C, Hu X, et al. (2022) K-lite: Learning transferable visual models with external knowledge. Advances in Neural Information Processing Systems 35:15558–15573
2022
Later among the works it cites.
Shi W, Shea R, Chen S, et al. (2022) Just fine-tune twice: Selective differential privacy for large language models. arXiv preprint arXiv:220407667
2022
Later among the works it cites.
Shreya G, Khapra MM (2022) A survey in adversarial defences and robustness in nlp. arXiv preprint arXiv:220306414
2022
Later among the works it cites.
Singh M, Gustafson L, Adcock A, et al. (2022) Revisiting weakly supervised pre-training of visual perception models. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 804–814
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xu L, Hu H, Zhang X, et al. (2020) CLUE: A Chinese language understanding evaluation benchmark. In: Proceedings of the 28th International Conference on Computational Linguistics. International Committee on Computational Linguistics, Barcelona, Spain (Online), pp 4762–4772, 10.18653/v1/2020.coling-main.419 , URL https://aclanthology.org/2020.coling-main.419
2020
Cited alongside, same era.
Yasunaga M, Liang P (2020) Graph-based, self-supervised program repair from diagnostic feedback. In: International Conference on Machine Learning, PMLR, pp 10799–10808
2020
Cited alongside, same era.
Zhong M, Liu P, Chen Y, et al. (2020) Extractive summarization as text matching. In: Jurafsky D, Chai J, Schluter N, et al. (eds) Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. Association for Computational Linguistics, Online, pp 6197–6208, 10.18653/v1/2020.acl-main.552 , URL https://aclanthology.org/2020.acl-main.552
2020
Cited alongside, same era.
Zhou J, Hu S, Lv X, et al. (2020) Kacc: A multi-task benchmark for knowledge abstraction, concretization and completion. arXiv preprint arXiv:200413631
2020
Cited alongside, same era.
Acquaviva S, Pu Y, Nye M, et al. (2021) Larc: Language annotated abstraction and reasoning corpus. In: Proceedings of the Annual Meeting of the Cognitive Science Society
2021
Cited alongside, same era.
Ahmad WU, Chakraborty S, Ray B, et al. (2021) Unified pre-training for program understanding and generation. arXiv preprint arXiv:210306333
2021
Cited alongside, same era.
Aroca-Ouellette S, Paik C, Roncone A, et al. (2021) Prost: Physical reasoning of objects through space and time. 2106.03634
2021
Cited alongside, same era.
Austin J, Odena A, Nye M, et al. (2021) Program synthesis with large language models. arXiv preprint arXiv:210807732
2021
Cited alongside, same era.
Balashankar A, Subramanian L (2021) Learning faithful representations of causal graphs. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers). Association for Computational Linguistics, Online, pp 839–850, 10.18653/v1/2021.acl-long.69 , URL https://aclanthology.org/2021.acl-long.69
2021
Cited alongside, same era.
2022
Later among the works it cites.
Stolfo A, Jin Z, Shridhar K, et al. (2022) A causal framework to quantify the robustness of mathematical reasoning with language models. arXiv preprint arXiv:221012023
2022
Later among the works it cites.
Sun J, Kousik S, Fridovich-Keil D, et al. (2022b) Self-supervised traffic advisors: Distributed, multi-view traffic prediction for smart cities. In: 2022 IEEE 25th International Conference on Intelligent Transportation Systems (ITSC), pp 917–922, 10.1109/ITSC55140.2022.9922340
2022
Later among the works it cites.
Tafjord O, Dalvi Mishra B, Clark P (2022) Entailer: Answering questions with faithful and truthful chains of reasoning. In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Abu Dhabi, United Arab Emirates, pp 2078–2093, 10.18653/v1/2022.emnlp-main.134 , URL https://aclanthology.org/2022.emnlp-main.134
2022
Later among the works it cites.
Tao C, Hou L, Zhang W, et al. (2022) Compression of generative pre-trained language models via quantization. 2203.10705
2022
Later among the works it cites.
Taylor R, Kardas M, Cucurull G, et al. (2022) Galactica: A large language model for science. arXiv preprint arXiv:221109085
2022
Later among the works it cites.
Team G (2022) GT4SD (Generative Toolkit for Scientific Discovery). URL https://github.com/GT4SD/gt4sd-core
2022
Later among the works it cites.
Thoppilan R, De Freitas D, Hall J, et al. (2022) Lamda: Language models for dialog applications. arXiv preprint arXiv:220108239
2022
Later among the works it cites.
Tian J, Li Y, Chen W, et al. (2022) Weakly supervised neural symbolic learning for cognitive tasks. In: Proceedings of the AAAI Conference on Artificial Intelligence, pp 5888–5896
2022
Later among the works it cites.
Tong Z, Song Y, Wang J, et al. (2022) Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training. 2203.12602
2022
Later among the works it cites.
Tsai HS, Chang HJ, Huang WC, et al. (2022) SUPERB-SG: Enhanced Speech processing Universal PERformance Benchmark for Semantic and Generative Capabilities. In: Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics, pp 8479–8492
2022
Later among the works it cites.
Tunstall L, Von Werra L, Wolf T (2022) Natural language processing with transformers. ” O’Reilly Media, Inc.”
2022
Later among the works it cites.
Valipour M, Rezagholizadeh M, Kobyzev I, et al. (2022) Dylora: Parameter efficient tuning of pre-trained models using dynamic search-free low-rank adaptation. arXiv preprint arXiv:221007558
2022
Later among the works it cites.
Welleck S, Lu X, West P, et al. (2022) Generating sequences by learning to self-correct. arXiv preprint arXiv:221100053
2022
Later among the works it cites.
Xiong J, Li C, Yang M, et al. (2022) Expression syntax information bottleneck for math word problems. In: Amigó E, Castells P, Gonzalo J, et al. (eds) SIGIR ’22: The 45th International ACM SIGIR Conference on Research and Development in Information Retrieval, Madrid, Spain, July 11 - 15, 2022. ACM, pp 2166–2171, 10.1145/3477495.3531824 , URL https://doi.org/10.1145/3477495.3531824
2022
Later among the works it cites.
Xu FF, Alon U, Neubig G, et al. (2022) A systematic evaluation of large language models of code. In: Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming, pp 1–10
2022
Later among the works it cites.
Xue F, Shi Z, Wei F, et al. (2022) Go wider instead of deeper. In: Proceedings of the AAAI Conference on Artificial Intelligence, pp 8779–8787
2022
Later among the works it cites.
Yang K, Deng J, Chen D (2022b) Generating natural language proofs with verifier-guided search. In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics, Abu Dhabi, United Arab Emirates, pp 89–105, 10.18653/v1/2022.emnlp-main.7 , URL https://aclanthology.org/2022.emnlp-main.7
2022
Later among the works it cites.
Yang Z, Yang Y (2022) Decoupling features in hierarchical propagation for video object segmentation. Advances in Neural Information Processing Systems 35:36324–36336
2022
Later among the works it cites.
Yao L, Han J, Wen Y, et al. (2022) Detclip: Dictionary-enriched visual-concept paralleled pre-training for open-world detection. Advances in Neural Information Processing Systems 35:9125–9138
2022
Later among the works it cites.
Ye X, Iyer S, Celikyilmaz A, et al. (2022) Complementary explanations for effective in-context learning. arXiv preprint arXiv:221113892
2022
Later among the works it cites.
Young N, Bao Q, Bensemann J, et al. (2022) AbductionRules: Training transformers to explain unexpected inputs. In: Findings of the Association for Computational Linguistics: ACL 2022. Association for Computational Linguistics, Dublin, Ireland, pp 218–227, 10.18653/v1/2022.findings-acl.19 , URL https://aclanthology.org/2022.findings-acl.19
2022
Later among the works it cites.
Yu S, Wu P, Liang PP, et al. (2022) Pacs: A dataset for physical audiovisual commonsense reasoning. 2203.11130
2022
Later among the works it cites.
Zan D, Chen B, Yang D, et al. (2022) Cert: Continual pre-training on sketches for library-oriented code generation. arXiv preprint arXiv:220606888
2022
Later among the works it cites.
Zelikman E, Wu Y, Mu J, et al. (2022) STar: Bootstrapping reasoning with reasoning. In: Oh AH, Agarwal A, Belgrave D, et al. (eds) Advances in Neural Information Processing Systems, URL https://openreview.net/forum?id=_3ELRdg2sgI
2022
Later among the works it cites.
Zeng A, Liu X, Du Z, et al. (2022) Glm-130b: An open bilingual pre-trained model. arXiv preprint arXiv:221002414
2022
Later among the works it cites.
Zhai X, Wang X, Mustafa B, et al. (2022) Lit: Zero-shot transfer with locked-image text tuning. In: CVPR, pp 18102–18112
2022
Later among the works it cites.
Zhang Y, Feng S, Tan C (2022b) Active example selection for in-context learning. In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, pp 9134–9148
2022
Later among the works it cites.
Zhao Y, Li Y, Li C, et al. (2022b) MultiHiertt: Numerical reasoning over multi hierarchical tabular and textual data. In: Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Dublin, Ireland, pp 6588–6600, 10.18653/v1/2022.acl-long.454 , URL https://aclanthology.org/2022.acl-long.454
2022
Later among the works it cites.
Zuo S, Zhang Q, Liang C, et al. (2022) Moebert: from bert to mixture-of-experts via importance-guided adaptation. arXiv preprint arXiv:220407675
2022
Later among the works it cites.
(2023) O (2023) Gpt-4 technical report. arXiv:230308774
2023
Closest in time.
Abdine H, Chatzianastasis M, Bouyioukos C, et al. (2023) Prot2text: Multimodal protein’s function generation with gnns and transformers. arXiv preprint arXiv:230714367
2023
Closest in time.
Aggarwal P, Madaan A, Yang Y, et al. (2023) Let’s sample step by step: Adaptive-consistency for efficient reasoning with llms. 2305.11860
2023
Closest in time.
Ai Q, Bai T, Cao Z, et al. (2023) Information retrieval meets large language models: A strategic report from chinese ir community. 2307.09751
2023
Closest in time.
Allal LB, Li R, Kocetkov D, et al. (2023) Santacoder: don’t reach for the stars! arXiv preprint arXiv:230103988
2023
Closest in time.
Anand Y, Nussbaum Z, Duderstadt B, et al. (2023) Gpt4all: Training an assistant-style chatbot with large scale data distillation from gpt-3.5-turbo. https://github.com/nomic-ai/gpt4all
2023
Closest in time.
Anil R, Dai AM, Firat O, et al. (2023) Palm 2 technical report. 2305.10403
2023
Closest in time.
Anthropic (2023) Introducing claude
2023
Closest in time.
Aryan A, Nain AK, McMahon A, et al. (2023) The costly dilemma: Generalization, evaluation and cost-optimal deployment of large language models. 2308.08061
2023
Closest in time.
Azerbayev Z, Schoelkopf H, Paster K, et al. (2023) Llemma: An open language model for mathematics. 2310.10631
2023
Closest in time.
Ban T, Chen L, Wang X, et al. (2023) From query tools to causal architects: Harnessing large language models for advanced causal discovery from data. arXiv preprint arXiv:230616902
2023
Closest in time.
Bao F, Nie S, Xue K, et al. (2023) One transformer fits all distributions in multi-modal diffusion at scale. 2303.06555
2023
Closest in time.
Berglund L, Tong M, Kaufmann M, et al. (2023) The reversal curse: Llms trained on” a is b” fail to learn” b is a”. arXiv preprint arXiv:230912288
2023
Closest in time.
Betker J, Goh G, Jing L, et al. (2023) Improving image generation with better captions. Computer Science https://cdn openai com/papers/dall-e-3 pdf
2023
Closest in time.
Bi K, Xie L, Zhang H, et al. (2023) Accurate medium-range global weather forecasting with 3d neural networks. Nature pp 1–6
2023
Closest in time.
Bowman SR (2023) Eight things to know about large language models. arXiv preprint arXiv:230400612
2023
Closest in time.
Brohan A, Brown N, Carbajal J, et al. (2023) Rt-2: Vision-language-action models transfer web knowledge to robotic control. In: TODO
2023
Closest in time.
Bubeck S, Chandrasekaran V, Eldan R, et al. (2023) Sparks of artificial general intelligence: Early experiments with gpt-4. 2303.12712
2023
Closest in time.
Bui ND, Le H, Wang Y, et al. (2023) Codetf: One-stop transformer library for state-of-the-art code llm. arXiv preprint arXiv:230600029
2023
Closest in time.
Cao Y, Xu X, Sun C, et al. (2023) Segment any anomaly without training via hybrid prompt regularization. arXiv preprint arXiv:230510724
2023
Closest in time.
Carlini N, Jagielski M, Choquette-Choo CA, et al. (2023) Poisoning web-scale training datasets is practical. arXiv preprint arXiv:230210149
2023
Closest in time.
Charalambous Y, Tihanyi N, Jain R, et al. (2023) A new era in software security: Towards self-healing software via large language models and formal verification. arXiv preprint arXiv:230514752
2023
Closest in time.
Cheng Y, Li L, Xu Y, et al. (2023) Segment and track anything. arXiv preprint arXiv:230506558
2023
Closest in time.
Cherti M, Beaumont R, Wightman R, et al. (2023) Reproducible scaling laws for contrastive language-image learning. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 2818–2829
2023
Closest in time.
Chiang WL, Li Z, Lin Z, et al. (2023) Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality. URL https://lmsys.org/blog/2023-03-30-vicuna/
2023
Closest in time.
Computer T (2023) Redpajama: An open source recipe to reproduce llama training dataset. URL https://github.com/togethercomputer/RedPajama-Data
2023
Closest in time.
Conover M, Hayes M, Mathur A, et al. (2023) Free dolly: Introducing the world’s first truly open instruction-tuned llm
2023
Closest in time.
Creswell A, Shanahan M, Higgins I (2023) Selection-inference: Exploiting large language models for interpretable logical reasoning. In: The Eleventh International Conference on Learning Representations, URL https://openreview.net/forum?id=3Pf3Wg6o-A4
2023
Closest in time.
Dai W, Liu Z, Ji Z, et al. (2023) Plausible may not be faithful: Probing object hallucination in vision-language pre-training. In: Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics. Association for Computational Linguistics, Dubrovnik, Croatia, pp 2136–2148, URL https://aclanthology.org/2023.eacl-main.156
2023
Closest in time.
Dettmers T, Pagnoni A, Holtzman A, et al. (2023) Qlora: Efficient finetuning of quantized llms. arXiv preprint arXiv:230514314
2023
Closest in time.
Di P, Li J, Yu H, et al. (2023) Codefuse-13b: A pretrained multi-lingual code large language model. arXiv preprint arXiv:231006266
2023
Closest in time.
2023
Closest in time.
DriveLM Contributors (2023) Drive on Language. URL https://github.com/OpenDriveLab/DriveLM/
2023
Closest in time.
Du Y, Li S, Torralba A, et al. (2023) Improving factuality and reasoning in language models through multiagent debate. 2305.14325
2023
Closest in time.
Echterhoff J, Yan A, Han K, et al. (2023) Driving through the concept gridlock: Unraveling explainability bottlenecks. arXiv preprint arXiv:231016639
2023
Closest in time.
Fan L, Krishnan D, Isola P, et al. (2023) Improving clip training with language rewrites. arXiv preprint arXiv:230520088
2023
Closest in time.
Firoozi R, Tucker J, Tian S, et al. (2023) Foundation models in robotics: Applications, challenges, and the future. arXiv preprint arXiv:231207843
2023
Closest in time.
Gadre SY, Ilharco G, Fang A, et al. (2023) Datacomp: In search of the next generation of multimodal datasets. arXiv preprint arXiv:230414108
2023
Closest in time.
Ge Y, Hua W, Mei K, et al. (2023) Openagi: When llm meets domain experts. 2304.04370
2023
Closest in time.
Gendron G, Bao Q, Witbrock M, et al. (2023) Large language models are not abstract reasoners. 2305.19555
2023
Closest in time.
Geng X, Gudibande A, Liu H, et al. (2023) Koala: A dialogue model for academic research. Blog post, URL https://bair.berkeley.edu/blog/2023/04/03/koala/
2023
Closest in time.
Gou Z, Shao Z, Gong Y, et al. (2023) Critic: Large language models can self-correct with tool-interactive critiquing. 2305.11738
2023
Closest in time.
gravitas/auto gpt S (2023) An experimental open-source attempt to make gpt-4 fully autonomou. 2305.16291
2023
Closest in time.
He H, Zhang J, Xu M, et al. (2023) Scalable mask annotation for video text spotting. arXiv preprint arXiv:230501443
2023
Closest in time.
Hong Y, Zhen H, Chen P, et al. (2023) 3d-llm: Injecting the 3d world into large language models. arXiv
2023
Closest in time.
Hsieh CY, Li CL, Yeh Ck, et al. (2023) Distilling step-by-step! outperforming larger language models with less training data and smaller model sizes. In: Findings of the Association for Computational Linguistics: ACL 2023. Association for Computational Linguistics, Toronto, Canada, pp 8003–8017, 10.18653/v1/2023.findings-acl.507 , URL https://aclanthology.org/2023.findings-acl.507
2023
Closest in time.
Imani S, Du L, Shrivastava H (2023) Mathprompter: Mathematical reasoning using large language models. 2303.05398
2023
Closest in time.
Inaba T, Kiyomaru H, Cheng F, et al. (2023) Multitool-cot: Gpt-3 can use multiple external tools with chain of thought prompting. 2305.16896
2023
Closest in time.
Jain N, Saifullah K, Wen Y, et al. (2023) Bring your own data! self-supervised evaluation for large language models. arXiv preprint arXiv:230613651
2023
Closest in time.
Ji Z, Lee N, Frieske R, et al. (2023) Survey of hallucination in natural language generation. ACM Computing Surveys 55(12):1–38
2023
Closest in time.
Kazemi M, Yuan Q, Bhatia D, et al. (2023) Boardgameqa: A dataset for natural language reasoning with contradictory information. arXiv preprint arXiv:230607934
2023
Closest in time.
Kıcıman E, Ness R, Sharma A, et al. (2023) Causal reasoning and large language models: Opening a new frontier for causality. arXiv preprint arXiv:230500050
2023
Closest in time.
Kim S, Joo SJ, Kim D, et al. (2023) The cot collection: Improving zero-shot and few-shot learning of language models via chain-of-thought fine-tuning. arXiv preprint arXiv:230514045
2023
Closest in time.
Kirchenbauer J, Geiping J, Wen Y, et al. (2023) A watermark for large language models. 2301.10226
2023
Closest in time.
Kirillov A, Mintun E, Ravi N, et al. (2023) Segment anything. arXiv:230402643
2023
Closest in time.
Kondo K, Sugawara S, Aizawa A (2023) Probing physical reasoning with counter-commonsense context. In: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers). Association for Computational Linguistics, Toronto, Canada, pp 603–612, 10.18653/v1/2023.acl-short.53 , URL https://aclanthology.org/2023.acl-short.53
2023
Closest in time.
Kosinski M (2023) Theory of mind may have spontaneously emerged in large language models. 2302.02083
2023
Closest in time.
Köksal A, Schick T, Korhonen A, et al. (2023) Longform: Optimizing instruction tuning for long text generation with corpus extraction. 2304.08460
2023
Closest in time.
Köpf A, Kilcher Y, von Rütte D, et al. (2023) Openassistant conversations – democratizing large language model alignment. 2304.07327
2023
Closest in time.
Laban P, Kryściński W, Agarwal D, et al. (2023) Llms as factual reasoners: Insights from existing benchmarks and beyond. 2305.14540
2023
Closest in time.
Lai X, Tian Z, Chen Y, et al. (2023) Lisa: Reasoning segmentation via large language model. arXiv preprint arXiv:230800692
2023
Closest in time.
Li C (2023) Large multimodal models: Notes on cvpr 2023 tutorial. 2306.14895
2023
Closest in time.
Li L, Spratling M (2023) Data augmentation alone can improve adversarial training. arXiv preprint arXiv:230109879
2023
Closest in time.
Li P, Sun T, Tang Q, et al. (2023j) Codeie: Large code generation models are better few-shot information extractors. In: Rogers A, Boyd-Graber JL, Okazaki N (eds) Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2023, Toronto, Canada, July 9-14, 2023. Association for Computational Linguistics, pp 15339–15353, 10.18653/v1/2023.acl-long.855 , URL https://doi.org/10.18653/v1/2023.acl-long.855
2023
Closest in time.
Li X, Lv K, Yan H, et al. (2023n) Unified demonstration retriever for in-context learning. In: Rogers A, Boyd-Graber J, Okazaki N (eds) Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Toronto, Canada, pp 4644–4668, 10.18653/v1/2023.acl-long.256 , URL https://aclanthology.org/2023.acl-long.256
2023
Closest in time.
Li Y, Gao C, Song X, et al. (2023r) Druggpt: A gpt-based strategy for designing potential ligands targeting specific proteins. bioRxiv pp 2023–06
2023
Closest in time.
Liang T, He Z, Jiao W, et al. (2023) Encouraging divergent thinking in large language models through multi-agent debate. arXiv preprint arXiv:230519118
2023
Closest in time.
Liao QV, Vaughan JW (2023) Ai transparency in the age of llms: A human-centered research roadmap. arXiv preprint arXiv:230601941
2023
Closest in time.
Lightman H, Kosaraju V, Burda Y, et al. (2023) Let’s verify step by step. 2305.20050
2023
Closest in time.
Long J (2023) Large language model guided tree-of-thought. 2305.08291
2023
Closest in time.
Long S, Piché A, Zantedeschi V, et al. (2023) Causal discovery with language models as imperfect experts. arXiv preprint arXiv:230702390
2023
Closest in time.
Longpre S, Hou L, Vu T, et al. (2023) The flan collection: Designing data and methods for effective instruction tuning. arXiv preprint arXiv:230113688
2023
Closest in time.
Lou R, Zhang K, Yin W (2023) Is prompt all you need? no. a comprehensive and broader view of instruction learning. arXiv preprint arXiv:230310475
2023
Closest in time.
Lu P, Peng B, Cheng H, et al. (2023) Chameleon: Plug-and-play compositional reasoning with large language models. 2304.09842
2023
Closest in time.
Ma X, Yong S, Zheng Z, et al. (2023) Sqa3d: Situated question answering in 3d scenes. In: International Conference on Learning Representations, URL https://openreview.net/forum?id=IDJx97BC38
2023
Closest in time.
Madaan A, Tandon N, Gupta P, et al. (2023) Self-refine: Iterative refinement with self-feedback. 2303.17651
2023
Closest in time.
Madani A, Krause B, Greene ER, et al. (2023) Large language models generate functional protein sequences across diverse families. Nature Biotechnology pp 1–8
2023
Closest in time.
Magister LC, Mallinson J, Adamek J, et al. (2023) Teaching small language models to reason. In: Rogers A, Boyd-Graber J, Okazaki N (eds) Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers). Association for Computational Linguistics, Toronto, Canada, pp 1773–1781, 10.18653/v1/2023.acl-short.151 , URL https://aclanthology.org/2023.acl-short.151
2023
Closest in time.
Manica M, Born J, Cadow J, et al. (2023) Accelerating material design with the generative toolkit for scientific discovery. npj Computational Materials 9(1):69
2023
Closest in time.
Mikuła M, Antoniak S, Tworkowski S, et al. (2023) Magnushammer: A transformer-based approach to premise selection. arXiv preprint arXiv:230304488
2023
Closest in time.
Mitra A, Del Corro L, Mahajan S, et al. (2023) Orca 2: Teaching small language models how to reason. arXiv preprint arXiv:231111045
2023
Closest in time.
Moghaddam SR, Honey CJ (2023) Boosting theory-of-mind performance in large language models via prompting. 2304.11490
2023
Closest in time.
Mu Y, Zhang Q, Hu M, et al. (2023) Embodiedgpt: Vision-language pre-training via embodied chain of thought. arXiv preprint arXiv:230515021
2023
Closest in time.
Mukherjee S, Mitra A, Jawahar G, et al. (2023) Orca: Progressive learning from complex explanation traces of gpt-4. arXiv preprint arXiv:230602707
2023
Closest in time.
Mündler N, He J, Jenko S, et al. (2023) Self-contradictory hallucinations of large language models: Evaluation, detection and mitigation. 2305.15852
2023
Closest in time.
Nascimento N, Alencar P, Cowan D (2023) Self-adaptive large language model (llm)-based multiagent systems. 2307.06187
2023
Closest in time.
Nguyen E, Poli M, Faizi M, et al. (2023) Hyenadna: Long-range genomic sequence modeling at single nucleotide resolution. arXiv preprint arXiv:230615794
2023
Closest in time.
Nijkamp E, Hayashi H, Xiong C, et al. (2023) Codegen2: Lessons for training llms on programming and natural languages. arXiv preprint arXiv:230502309
2023
Closest in time.
Ning X, Lin Z, Zhou Z, et al. (2023) Skeleton-of-thought: Large language models can do parallel decoding. 2307.15337
2023
Closest in time.
Olausson TX, Inala JP, Wang C, et al. (2023) Demystifying gpt self-repair for code generation. 2306.09896
2023
Closest in time.
Padalkar A, Pooley A, Jain A, et al. (2023) Open x-embodiment: Robotic learning datasets and rt-x models. arXiv preprint arXiv:231008864
2023
Closest in time.
Paranjape B, Lundberg S, Singh S, et al. (2023) Art: Automatic multi-step reasoning and tool-use for large language models. 2303.09014
2023
Closest in time.
Paster K, Santos MD, Azerbayev Z, et al. (2023) Openwebmath: An open dataset of high-quality mathematical web text. 2310.06786
2023
Closest in time.
Paul D, Ismayilzada M, Peyrard M, et al. (2023) Refiner: Reasoning feedback on intermediate representations. 2304.01904
2023
Closest in time.
Pham H, Dai Z, Ghiasi G, et al. (2023) Combined scaling for zero-shot transfer learning. Neurocomputing 555:126658
2023
Closest in time.
Pi R, Gao J, Diao S, et al. (2023) Detgpt: Detect what you need via reasoning. 2305.14167
2023
Closest in time.
Poli M, Massaroli S, Nguyen E, et al. (2023) Hyena hierarchy: Towards larger convolutional language models. arXiv preprint arXiv:230210866
2023
Closest in time.
Polu S, Han JM, Zheng K, et al. (2023) Formal mathematics statement curriculum learning. In: The Eleventh International Conference on Learning Representations
2023
Closest in time.
Press O, Zhang M, Min S, et al. (2023) Measuring and narrowing the compositionality gap in language models. 2210.03350
2023
Closest in time.
Pryor C, Dickens C, Augustine E, et al. (2023) Neupsl: Neural probabilistic soft logic. 2205.14268
2023
Closest in time.
Qiao S, Gui H, Chen H, et al. (2023) Making language models better tool learners with execution feedback. 2305.13068
2023
Closest in time.
Qin Y, Hu S, Lin Y, et al. (2023) Tool learning with foundation models. 2304.08354
2023
Closest in time.
Rafailov R, Sharma A, Mitchell E, et al. (2023) Direct preference optimization: Your language model is secretly a reward model. arXiv preprint arXiv:230518290
2023
Closest in time.
2023
Closest in time.
Rawte V, Sheth A, Das A (2023) A survey of hallucination in large foundation models. 2309.05922
2023
Closest in time.
Ren S, Zhu KQ (2023) Low-rank prune-and-factorize for language model compression. 2306.14152
2023
Closest in time.
Ren X, Zhou P, Meng X, et al. (2023) Pangu- \ s i g m a \backslash sigma : Towards trillion parameter language model with sparse heterogeneous computing. arXiv preprint arXiv:230310845
2023
Closest in time.
Roziere B, Gehring J, Gloeckle F, et al. (2023) Code llama: Open foundation models for code. arXiv preprint arXiv:230812950
2023
Closest in time.
Saparov A, He H (2023) Language models are greedy reasoners: A systematic formal analysis of chain-of-thought. In: The Eleventh International Conference on Learning Representations, URL https://openreview.net/forum?id=qFVVBzXxR2V
2023
Closest in time.
Savage N (2023) Drug discovery companies are customizing chatgpt: here’s how. Nature Biotechnology
2023
Closest in time.
Sawada T, Paleka D, Havrilla A, et al. (2023) Arb: Advanced reasoning benchmark for large language models. 2307.13692
2023
Closest in time.
Schick T, Dwivedi-Yu J, Dessì R, et al. (2023) Toolformer: Language models can teach themselves to use tools. 2302.04761
2023
Closest in time.
Seff A, Cera B, Chen D, et al. (2023) MotionLM: Multi-Agent Motion Forecasting as Language Modeling. In: ICCV
2023
Closest in time.
Sha H, Mu Y, Jiang Y, et al. (2023) Languagempc: Large language models as decision makers for autonomous driving. arXiv preprint arXiv:231003026
2023
Closest in time.
Shao Z, Yu Z, Wang M, et al. (2023) Prompting large language models with answer heuristics for knowledge-based visual question answering. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 14974–14983
2023
Closest in time.
Shen Y, Song K, Tan X, et al. (2023) Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face. 2303.17580
2023
Closest in time.
Shi F, Suzgun M, Freitag M, et al. (2023) Language models are multilingual chain-of-thought reasoners. In: The Eleventh International Conference on Learning Representations, URL https://openreview.net/forum?id=fR3wGCk-IXp
2023
Closest in time.
Shinn N, Cassano F, Labash B, et al. (2023) Reflexion: Language agents with verbal reinforcement learning. 2303.11366
2023
Closest in time.
Shridhar K, Stolfo A, Sachan M (2023) Distilling reasoning capabilities into smaller language models. In: Findings of the Association for Computational Linguistics: ACL 2023. Association for Computational Linguistics, Toronto, Canada, pp 7059–7073, 10.18653/v1/2023.findings-acl.441 , URL https://aclanthology.org/2023.findings-acl.441
2023
Closest in time.
Singh I, Blukis V, Mousavian A, et al. (2023) Progprompt: Generating situated robot task plans using large language models. In: 2023 IEEE International Conference on Robotics and Automation (ICRA), IEEE, pp 11523–11530
2023
Closest in time.
Singhal K, Tu T, Gottweis J, et al. (2023) Towards expert-level medical question answering with large language models. arXiv preprint arXiv:230509617
2023
Closest in time.
Soldaini L, Lo K (2023) peS2o (Pretraining Efficiently on S2ORC) Dataset. Tech. rep., Allen Institute for AI, oDC-By, https://github.com/allenai/pes2o
2023
Closest in time.
Srivastava A, Rastogi A, Rao A, et al. (2023) Beyond the imitation game: Quantifying and extrapolating the capabilities of language models. Transactions on Machine Learning Research URL https://openreview.net/forum?id=uyTL5Bvosj
2023
Closest in time.
Subramanian S, Harrington P, Keutzer K, et al. (2023) Towards foundation models for scientific machine learning: Characterizing scaling and transfer behavior. 2306.00258
2023
Closest in time.
Tan S, Ivanovic B, Weng X, et al. (2023) Language conditioned traffic generation. CoRL
2023
Closest in time.
Tao M, Bao BK, Tang H, et al. (2023) Galip: Generative adversarial clips for text-to-image synthesis. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 14214–14223
2023
Closest in time.
Taori R, Gulrajani I, Zhang T, et al. (2023) Alpaca: A strong, replicable instruction-following model. Stanford Center for Research on Foundation Models https://crfm stanford edu/2023/03/13/alpaca html 3(6):7
2023
Closest in time.
Team G, Anil R, Borgeaud S, et al. (2023) Gemini: a family of highly capable multimodal models. arXiv preprint arXiv:231211805
2023
Closest in time.
Tung HY, Ding M, Chen Z, et al. (2023) Physion++: Evaluating physical scene understanding that requires online inference of different physical properties. arXiv preprint arXiv:230615668
2023
Closest in time.
Wan Z, Cheng F, Mao Z, et al. (2023) Gpt-re: In-context learning for relation extraction using large language models. 2305.02105
2023
Closest in time.
Wang X, Gu R, Chen Z, et al. (2023r) Uni-rna: universal pre-trained models revolutionize rna research. bioRxiv pp 2023–07
2023
Closest in time.
Watson JL, Juergens D, Bennett NR, et al. (2023) De novo design of protein structure and function with rfdiffusion. Nature 620(7976):1089–1100
2023
Closest in time.
Weston J, Sukhbaatar S (2023) System 2 attention (is something you might need too). 2311.11829
2023
Closest in time.
Willig M, Zečević M, Dhami DS, et al. (2023) Probing for correlations of causal facts: Large language models and causality. URL https://openreview.net/forum?id=UPwzqPOs4-
2023
Closest in time.
Xi Z, Chen W, Guo X, et al. (2023) The rise and potential of large language model based agents: A survey. 2309.07864
2023
Closest in time.
Xin H, Wang H, Zheng C, et al. (2023) Lego-prover: Neural theorem proving with growing libraries. arXiv preprint arXiv:231000656
2023
Closest in time.
Yan Z, Zhang K, Zhou R, et al. (2023) Multimodal chatgpt for medical applications: an experimental study of gpt-4v. arXiv preprint arXiv:231019061
2023
Closest in time.
Yang H, Wang Y, Li P, et al. (2023a) Bridging the gap between pre-training and fine-tuning for commonsense generation. In: Findings of the Association for Computational Linguistics: EACL 2023, pp 376–383
2023
Closest in time.
Yin Z, Sun Q, Chang C, et al. (2023c) Exchange-of-thought: Enhancing large language model capabilities through cross-model communication. In: The 2023 Conference on Empirical Methods in Natural Language Processing, URL https://openreview.net/forum?id=30kbnyD9hF
2023
Closest in time.
Yin Z, Sun Q, Guo Q, et al. (2023d) Do large language models know what they don’t know? In: Rogers A, Boyd-Graber J, Okazaki N (eds) Findings of the Association for Computational Linguistics: ACL 2023. Association for Computational Linguistics, Toronto, Canada, pp 8653–8665, 10.18653/v1/2023.findings-acl.551 , URL https://aclanthology.org/2023.findings-acl.551
2023
Closest in time.
Yoneda T, Fang J, Li P, et al. (2023) Statler: State-maintaining language models for embodied reasoning. 2306.17840
2023
Closest in time.
Yue X, Qu X, Zhang G, et al. (2023) Mammoth: Building math generalist models through hybrid instruction tuning. arXiv preprint arXiv:230905653
2023
Closest in time.
Zan D, Chen B, Zhang F, et al. (2023) Large language models meet nl2code: A survey. In: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp 7443–7464
2023
Closest in time.
Zečević M, Willig M, Dhami DS, et al. (2023) Causal parrots: Large language models may talk causality but are not causal. arXiv preprint arXiv:230813067
2023
Closest in time.
Zeng A, Attarian M, brian ichter, et al. (2023) Socratic models: Composing zero-shot multimodal reasoning with language. In: The Eleventh International Conference on Learning Representations, URL https://openreview.net/forum?id=G2Q2Mh3avow
2023
Closest in time.
Zhang K, Li Z, Li J, et al. (2023e) Self-edit: Fault-aware code editor for code generation. In: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). Association for Computational Linguistics, Toronto, Canada, pp 769–787, URL https://aclanthology.org/2023.acl-long.45
2023
Closest in time.
Zhong T, Zhao W, Zhang Y, et al. (2023) Chatradio-valuer: A chat large language model for generalizable radiology report generation based on multi-institution and multi-system data. arXiv preprint arXiv:231005242
2023
Closest in time.
Zhou J, Zhang Y, Luo Q, et al. (2023b) Synthetic lies: Understanding ai-generated misinformation and evaluating algorithmic and human solutions. In: Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems, pp 1–20
2023
Closest in time.
Ziyu Z, Xiaojian M, Yixin C, et al. (2023) 3d-vista: Pre-trained transformer for 3d vision and text alignment. In: ICCV
2023
Closest in time.
Zong Y, Aodha OM, Hospedales T (2023) Self-supervised multimodal learning: A survey. 2304.01008
2023
Closest in time.
Yih Wt, Richardson M, Meek C, et al. (2016) The value of semantic parse labeling for knowledge base question answering. In: Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers). Association for Computational Linguistics, Berlin, Germany, pp 201–206, 10.18653/v1/P16-2033 , URL https://aclanthology.org/P16-2033
2033
Closest in time.
Nunes T (2012) Logical Reasoning and Learning, Springer US, Boston, MA, pp 2066–2069. 10.1007/978-1-4419-1428-6_790 , URL https://doi.org/10.1007/978-1-4419-1428-6_790
2069
Closest in time.