Fetching the paper…
Reading the bibliography…
The process of identifying changes or transformations in a scene along with the ability of reasoning about their causes and effects, is a key aspect of intelligence.
Temporal relational reasoning in videos
Bolei Zhou, Alex Andonian, Aude Oliva, and Antonio Torralba · 1905
Earlier work this paper cites.
Play, Dreams, and Imitation in Childhood
Jean Piaget · 1962
Earlier work this paper cites.
The Art of Block Building
Harriet Johnson · 1983
Earlier work this paper cites.
Play can be the building blocks of learning
Sally S. Cartwright · 1988
Earlier work this paper cites.
Inductive logic programming
Stephen Muggleton · 1991
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, Patrick Haffner, et al · 1998
Earlier work this paper cites.
Automatic analysis of the difference image for unsupervised change detection
Lorenzo Bruzzone and Diego F Prieto · 2000
Earlier work this paper cites.
Building blocks and cognitive building blocks - playing to know the world mathematically
Julie S. Sarama and Douglas H. Clements · 2001
Cited alongside, same era.
Blocks world revisited: Image understanding using qualitative geometry and mechanics
Abhinav Gupta, Alexei A Efros, and Martial Hebert · 2010
Cited alongside, same era.
Potassco: The potsdam answer set solving collection
Martin Gebser, Benjamin Kaufmann, Roland Kaminski, Max Ostrowski, Torsten Schaub, and Marius Schneider · 2011
Cited alongside, same era.
Large-scale damage detection using satellite imagery
Lionel Gueguen and Raffay Hamid · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Addressing a question answering challenge by combining statistical methods with inductive rule learning and reasoning
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Later among the works it cites.
Phrase localization and visual relationship detection with comprehensive image-language cues
Bryan A Plummer, Arun Mallya, Christopher M Cervantes, Julia Hockenmaier, and Svetlana Lazebnik · 2017
Later among the works it cites.
A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David G Barrett, Mateusz Malinowski, Razvan Pascanu, Peter Battaglia, and Timothy Lillicrap · 2017
Later among the works it cites.
Pyramid scene parsing network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia · 2017
Later among the works it cites.
Encoder-decoder with atrous separable convolution for semantic image segmentation
Liang-Chieh Chen, Yukun Zhu, George Papandreou, Florian Schroff, and Hartwig Adam · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Arindam Mitra and Chitta Baral · 2016
Cited alongside, same era.
You only look once: Unified, real-time object detection
Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi · 2016
Cited alongside, same era.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Cited alongside, same era.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Cited alongside, same era.
Harsh Jhamtani and Taylor Berg-Kirkpatrick · 2018
Later among the works it cites.
Large-scale visual relationship understanding
Ji Zhang, Yannis Kalantidis, Marcus Rohrbach, Manohar Paluri, Ahmed M. Elgammal, and Mohamed Elhoseiny · 2018
Later among the works it cites.
Modularized textual grounding for counterfactual resilience
Zhiyuan Fang, Shu Kong, Charless Fowlkes, and Yezhou Yang · 2019
Closest in time.
Viewpoint invariant change captioning
Dong Huk Park, Trevor Darrell, and Anna Rohrbach · 2019
Closest in time.