Fetching the paper…
Reading the bibliography…
Recent times have witnessed an increasing number of applications of deep neural networks towards solving tasks that require superior cognitive abilities, e.g., playing Go, generating art, ChatGPT, etc.
Steps toward artificial intelligence
Marvin Minsky · 1961
Earlier work this paper cites.
A program for the solution of a class of geometric-analogy intelligence-test questions
Thomas G Evans · 1964
Earlier work this paper cites.
Society of mind
Marvin Minsky · 1988
Earlier work this paper cites.
Fluid concepts and creative analogies: Computer models of the fundamental mechanisms of thought
Douglas R Hofstadter · 1995
Earlier work this paper cites.
A theory of causal learning in children: causal maps and bayes nets
Alison Gopnik, Clark Glymour, David M Sobel, Laura E Schulz, Tamar Kushnir, and David Danks · 2004
Earlier work this paper cites.
A new framework for understanding how young children create external representations for puzzles and problems
Lara M Triona and David Klahr · 2007
Earlier work this paper cites.
Exploring network structure, dynamics, and function using networkx
Aric A. Hagberg, Daniel A. Schult, and Pieter J. Swart · 2008
Earlier work this paper cites.
Infants consider both the sample and the sampling process in inductive generalization
Hyowon Gweon, Joshua B Tenenbaum, and Laura E Schulz · 2010
Earlier work this paper cites.
Where science starts: Spontaneous experiments in preschoolers’ exploratory play
Claire Cook, Noah D Goodman, and Laura E Schulz · 2011
Earlier work this paper cites.
Supporting inquiry about the foundations of evolutionary thinking in the elementary grades
Richard Lehrer and Leona Schauble · 2012
Earlier work this paper cites.
Graphic symbols as “the mind on paper”: Links between children’s interpretive theory of mind and symbol understanding
Lauren J Myers and Lynn S Liben · 2012
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Neural module networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Earlier work this paper cites.
Computer models solving intelligence test problems: Progress and implications
José Hernández-Orallo, Fernando Martínez-Plumed, Ute Schmid, Michael Siebers, and David L Dowe · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens Van Der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Earlier work this paper cites.
DeepStory: Video story QA by deep embedded memory networks
Kyung-Min Kim, Min-Oh Heo, Seong-Ho Choi, and Byoung-Tak Zhang · 2017
Earlier work this paper cites.
Building machines that learn and think like people
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman · 2017
Earlier work this paper cites.
Combining knowledge and reasoning through probabilistic soft logic for image puzzle solving
Somak Aditya, Yezhou Yang, Chitta Baral, and Yiannis Aloimonos · 2018
Earlier work this paper cites.
Measuring abstract reasoning in neural networks
David G.T. Barrett, Felix Hill, Adam Santoro, Ari S. Morcos, and Timothy Lillicrap · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Benchmarking neural network robustness to common corruptions and surface variations
Dan Hendrycks and Thomas G Dietterich · 2018
Cited alongside, same era.
Visual entailment task for visually-grounded language learning
Ning Xie, Farley Lai, Derek Doran, and Asim Kadav · 2018
Cited alongside, same era.
On the measure of intelligence
François Chollet · 2019
Cited alongside, same era.
Learning by abstraction: The neural state machine
Drew Hudson and Christopher D Manning · 2019
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Cited alongside, same era.
Exploring simple siamese representation learning
Xinlei Chen and Kaiming He · 2021
Later among the works it cites.
Stratified rule-aware network for abstract visual reasoning
Sheng Hu, Yuqing Ma, Xianglong Liu, Yanlu Wei, and Shihao Bai · 2021
Later among the works it cites.
Adversarial VQA: A new benchmark for evaluating the robustness of VQA models
Linjie Li, Jie Lei, Zhe Gan, and Jingjing Liu · 2021
Later among the works it cites.
Swin Transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Later among the works it cites.
Abstraction and analogy-making in artificial intelligence
Melanie Mitchell · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Cited alongside, same era.
Explainable and explicit visual reasoning over scene graphs
Jiaxin Shi, Hanwang Zhang, and Juanzi Li · 2019
Cited alongside, same era.
Towards vqa models that can read
Amanpreet Singh, Vivek Natarajan, Meet Shah, Yu Jiang, Xinlei Chen, Dhruv Batra, Devi Parikh, and Marcus Rohrbach · 2019
Cited alongside, same era.
A corpus for reasoning about natural language grounded in photographs
Alane Suhr, Stephanie Zhou, Ally Zhang, Iris Zhang, Huajun Bai, and Yoav Artzi · 2019
Cited alongside, same era.
From recognition to cognition: Visual commonsense reasoning
Rowan Zellers, Yonatan Bisk, Ali Farhadi, and Yejin Choi · 2019
Cited alongside, same era.
RAVEN: A dataset for relational and analogical visual reasoning
Chi Zhang, Feng Gao, Baoxiong Jia, Yixin Zhu, and Song-Chun Zhu · 2019
Cited alongside, same era.
The gap of semantic parsing: A survey on automatic math word problem solvers
Dongxiang Zhang, Lei Wang, Luming Zhang, Bing Tian Dai, and Heng Tao Shen · 2019
Cited alongside, same era.
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Later among the works it cites.
Human-adversarial visual question answering
Sasha Sheng, Amanpreet Singh, Vedanuj Goswami, Jose Alberto Lopez Magana, Wojciech Galuba, Devi Parikh, and Douwe Kiela · 2021
Later among the works it cites.
Using program synthesis and inductive logic programming to solve bongard problems
Atharv Sonwane, Sharad Chitlangia, Tirtharaj Dash, Lovekesh Vig, Gautam Shroff, and Ashwin Srinivasan · 2021
Later among the works it cites.
Unsupervised abstract reasoning for raven’s problem matrices
Tao Zhuo, Qiang Huang, and Mohan Kankanhalli · 2021
Later among the works it cites.
https://mathkangaroo.org/mks/, 2012–2022
Math Kangaroo USA, NFP Inc · 2022
Closest in time.
WinoGAViL: Gamified association benchmark to challenge vision-and-language models
Yonatan Bitton, Nitzan Bitton Guetta, Ron Yosef, Yuval Elovici, Mohit Bansal, Gabriel Stanovsky, and Roy Schwartz · 2022
Closest in time.
A neural network solves, explains, and generates university math problems by program synthesis and few-shot learning at human level
Iddo Drori, Sarah Zhang, Reece Shuttleworth, Leonard Tang, Albert Lu, Elizabeth Ke, Kevin Liu, Linda Chen, Sunny Tran, Newman Cheng, et al · 2022
Closest in time.
Discovering faster matrix multiplication algorithms with reinforcement learning
Alhussein Fawzi, Matej Balog, Aja Huang, Thomas Hubert, Bernardino Romera-Paredes, Mohammadamin Barekatain, Alexander Novikov, Francisco J R Ruiz, Julian Schrittwieser, Grzegorz Swirszcz, et al · 2022
Closest in time.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Closest in time.
A neuro-vector-symbolic architecture for solving raven’s progressive matrices
Michael Hersche, Mustafa Zeqiri, Luca Benini, Abu Sebastian, and Abbas Rahimi · 2022
Closest in time.
Bongard-HOI: Benchmarking few-shot visual reasoning for human-object interactions
Huaizu Jiang, Xiaojian Ma, Weili Nie, Zhiding Yu, Yuke Zhu, and Anima Anandkumar · 2022
Closest in time.
A review of emerging research directions in abstract visual reasoning
Mikołaj Małkiński and Jacek Mańdziuk · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Closest in time.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S Sara Mahdavi, Rapha Gontijo Lopes, et al · 2022
Closest in time.
Flava: A foundational language and vision alignment model
Amanpreet Singh, Ronghang Hu, Vedanuj Goswami, Guillaume Couairon, Wojciech Galuba, Marcus Rohrbach, and Douwe Kiela · 2022
Closest in time.
Winoground: Probing vision and language models for visio-linguistic compositionality
Tristan Thrush, Ryan Jiang, Max Bartolo, Amanpreet Singh, Adina Williams, Douwe Kiela, and Candace Ross · 2022
Closest in time.
CrossFormer: A versatile vision transformer hinging on cross-scale attention
Wenxiao Wang, Lu Yao, Long Chen, Binbin Lin, Deng Cai, Xiaofei He, and Wei Liu · 2022
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.