Fetching the paper…
Reading the bibliography…
Relational reasoning is a central component of generally intelligent behavior, but has proven difficult for neural networks to learn.
Physical symbol systems
Allen Newell · 1980
Earlier work this paper cites.
The symbol grounding problem
Stevan Harnad · 1990
Earlier work this paper cites.
A computational analysis of the apprehension of spatial relations
Gordon D Logan and Daniel D Sadler · 1996
Earlier work this paper cites.
Learning to parse database queries using inductive logic programming
John M Zelle and Raymond J Mooney · 1996
Earlier work this paper cites.
Situated dialogue and spatial organization: What, where… and why
Geert-Jan M Kruijff, Hendrik Zender, Patric Jensfelt, and Henrik I Christensen · 2007
Earlier work this paper cites.
The discovery of structural form
Charles Kemp and Joshua B Tenenbaum · 2008
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini · 2009
Earlier work this paper cites.
A game-theoretic approach to generating spatial descriptions
Dave Golland, Percy Liang, and Dan Klein · 2010
Earlier work this paper cites.
Inducing probabilistic ccg grammars from logical form with higher-order unification
Tom Kwiatkowski, Luke Zettlemoyer, Sharon Goldwater, and Mark Steedman · 2010
Earlier work this paper cites.
Grounding spatial language for video search
Stefanie Tellex, Thomas Kollar, George Shaw, Nicholas Roy, and Deb Roy · 2010
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
Stefanie Tellex, Thomas Kollar, Steven Dickerson, Matthew R Walter, Ashis Gopal Banerjee, Seth J Teller, and Nicholas Roy · 2011
Earlier work this paper cites.
Image retrieval with structured object queries using latent ranking svm
Tian Lan, Weilong Yang, Yang Wang, and Greg Mori · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Semantic parsing on freebase from question-answer pairs
Jonathan Berant, Andrew Chou, Roy Frostig, and Percy Liang · 2013
Earlier work this paper cites.
Grounding spatial relations for human-robot interaction
Sergio Guadarrama, Lorenzo Riano, Dave Golland, Daniel Gouhring, Yangqing Jia, Dan Klein, Pieter Abbeel, and Trevor Darrell · 2013
Earlier work this paper cites.
Jointly learning to parse and perceive: Connecting natural language to the physical world
Jayant Krishnamurthy and Thomas Kollar · 2013
Earlier work this paper cites.
Learning dependency-based compositional semantics
Percy Liang, Michael I Jordan, and Dan Klein · 2013
Cited alongside, same era.
A multi-world approach to question answering about real-world scenes based on uncertain input
Mateusz Malinowski and Mario Fritz · 2014
Cited alongside, same era.
A pooling approach to modelling spatial relations for image retrieval and annotation
Mateusz Malinowski and Mario Fritz · 2014
Cited alongside, same era.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Cited alongside, same era.
Are you talking to a machine? dataset and methods for multilingual image question answering
Haoyuan Gao, Junhua Mao, Jie Zhou, Zhiheng Huang, Lei Wang, and Wei Xu · 2015
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling · 2016
Later among the works it cites.
Building machines that learn and think like people
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman · 2016
Later among the works it cites.
Gated graph sequence neural networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard Zemel · 2016
Later among the works it cites.
Ask your neurons: A deep learning approach to visual question answering
Mateusz Malinowski, Marcus Rohrbach, and Mario Fritz · 2016
Later among the works it cites.
Learning convolutional neural networks for graphs
Mathias Niepert, Mohamed Ahmed, and Konstantin Kutzkov · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mikael Henaff, Joan Bruna, and Yann LeCun · 2015
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Cited alongside, same era.
Image question answering: A visual semantic embedding model and a new dataset
Mengye Ren, Ryan Kiros, and Richard Zemel · 2015
Cited alongside, same era.
Towards ai-complete question answering: A set of prerequisite toy tasks
Jason Weston, Antoine Bordes, Sumit Chopra, and Tomas Mikolov · 2015
Cited alongside, same era.
Memory networks
Jason Weston, Sumit Chopra, and Antoine Bordes · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Cited alongside, same era.
Interaction networks for learning about objects, relations and physics
Peter Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, et al · 2016
Cited alongside, same era.
Scaling memory-augmented neural networks with sparse reads and writes
Jack Rae, Jonathan J Hunt, Ivo Danihelka, Timothy Harley, Andrew W Senior, Gregory Wayne, Alex Graves, and Tim Lillicrap · 2016
Later among the works it cites.
Dynamic memory networks for visual and textual question answering
Caiming Xiong, Stephen Merity, and Richard Socher · 2016
Later among the works it cites.
Ask, attend and answer: Exploring question-guided spatial attention for visual question answering
Huijuan Xu and Kate Saenko · 2016
Later among the works it cites.
Stacked attention networks for image question answering
Zichao Yang, Xiaodong He, Jianfeng Gao, Li Deng, and Alex Smola · 2016
Later among the works it cites.
Tracking the world state with recurrent entity networks
Mikael Henaff, Jason Weston, Arthur Szlam, Antoine Bordes, and Yann LeCun · 2017
Closest in time.
Learning to reason: End-to-end module networks for visual question answering
Ronghang Hu, Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Kate Saenko · 2017
Closest in time.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Closest in time.
Inferring and executing programs for visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Judy Hoffman, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Closest in time.
An analysis of visual question answering algorithms
Kushal Kafle and Christopher Kanan · 2017
Closest in time.
Discovering objects and their relations from entangled scene representations
David Raposo, Adam Santoro, David Barrett, Razvan Pascanu, Timothy Lillicrap, and Peter Battaglia · 2017
Closest in time.