Fetching the paper…
Reading the bibliography…
A fundamental component of human vision is our ability to parse complex visual scenes and judge the relations between their constituent objects.
The recognition problem
Mikhail Moiseevich Bongard · 1968
Earlier work this paper cites.
Raven’s progressive matrices (1938): More on norms, reliability, and validity
Henry R Burke · 1985
Earlier work this paper cites.
Visual routines
Shimon Ullman · 1987
Earlier work this paper cites.
Visual cognition
Patrick Cavanagh · 2011
Earlier work this paper cites.
Comparing machines and humans on a visual categorization test
François Fleuret, Ting Li, Charles Dubout, Emma K Wampler, Steven Yantis, and Donald Geman · 2011
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Ian J Goodfellow, Jonathon Shlens, and Christian Szegedy · 2014
Earlier work this paper cites.
Deep residual learning for image recognition." computer vision and pattern recognition (2015)
Kaiming He, X Zhang, S Ren, and J Sun · 2015
Earlier work this paper cites.
Human-level concept learning through probabilistic program induction
Brenden M Lake, Ruslan Salakhutdinov, and Joshua B Tenenbaum · 2015
Earlier work this paper cites.
Automatic generation of raven’s progressive matrices
Ke Wang and Zhendong Su · 2015
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens Van Der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Earlier work this paper cites.
A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David G Barrett, Mateusz Malinowski, Razvan Pascanu, Peter Battaglia, and Timothy Lillicrap · 2017
Earlier work this paper cites.
Measuring abstract reasoning in neural networks
David Barrett, Felix Hill, Adam Santoro, Ari Morcos, and Timothy Lillicrap · 2018
Cited alongside, same era.
Not-so-clevr: learning same–different relations strains feedforward neural networks
Junkyung Kim, Matthew Ricci, and Thomas Serre · 2018
Cited alongside, same era.
Learning long-range spatial dependencies with horizontal gated recurrent units
Drew Linsley, Junkyung Kim, Vijay Veerabadran, Charles Windolf, and Thomas Serre · 2018
Cited alongside, same era.
Phyre: A new benchmark for physical reasoning
Anton Bakhtin, Laurens van der Maaten, Justin Johnson, Laura Gustafson, and Ross Girshick · 2019
Cited alongside, same era.
On the measure of intelligence
François Chollet · 2019
Cited alongside, same era.
Recurrent neural circuits for contour detection
Drew Linsley, Junkyung Kim, Alekh Ashok, and Thomas Serre · 2020
Later among the works it cites.
Bongard-LOGO: A new benchmark for human-level concept learning and reasoning
Weili Nie, Zhiding Yu, Lei Mao, Ankit B Patel, Yuke Zhu, and Anima Anandkumar · 2020
Later among the works it cites.
V-prom: A benchmark for visual reasoning using visual progressive matrices
Damien Teney, Peng Wang, Jiewei Cao, Lingqiao Liu, Chunhua Shen, and Anton van den Hengel · 2020
Later among the works it cites.
Yuhuai Wu, Honghua Dong, Roger Grosse, and Jimmy Ba · 2020
Later among the works it cites.
An empirical study of training self-supervised vision transformers
Xinlei Chen, Saining Xie, and Kaiming He · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Junkyung Kim, Drew Linsley, Kalpit Thakkar, and Thomas Serre · 2019
Cited alongside, same era.
Clevrer: Collision events for video representation and reasoning
Kexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli, Jiajun Wu, Antonio Torralba, and Joshua B Tenenbaum · 2019
Cited alongside, same era.
Raven: A dataset for relational and analogical visual reasoning
Chi Zhang, Feng Gao, Baoxiong Jia, Yixin Zhu, and Song-Chun Zhu · 2019
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Shortcut learning in deep neural networks
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge, and Felix A Wichmann · 2020
Cited alongside, same era.
Recurrent vision transformer for solving visual reasoning problems
Nicola Messina, Giuseppe Amato, Fabio Carrara, Claudio Gennaro, and Fabrizio Falchi · 2021
Later among the works it cites.
Can deep convolutional neural networks support relational reasoning in the same-different task?
Guillermo Puebla and Jeffrey S Bowers · 2021
Later among the works it cites.
Bongard-hoi: Benchmarking few-shot visual reasoning for human-object interactions
Huaizu Jiang, Xiaojian Ma, Weili Nie, Zhiding Yu, Yuke Zhu, and Anima Anandkumar · 2022
Closest in time.
Qlevr: A diagnostic dataset for quantificational language and elementary visual reasoning
Zechen Li and Anders Søgaard · 2022
Closest in time.
Understanding the computational demands underlying visual reasoning
Mohit Vaishnav, Remi Cadene, Andrea Alamia, Drew Linsley, Rufin VanRullen, and Thomas Serre · 2022
Closest in time.