Fetching the paper…
Reading the bibliography…
Learning long-term dynamics models is the key to understanding physical common sense.
The nature of explanation
Kenneth James Williams Craik · 1952
Earlier work this paper cites.
Intuitive physics
Michael McCloskey · 1983
Earlier work this paper cites.
Forward models: Supervised learning with a distal teacher
Michael I Jordan and David E Rumelhart · 1992
Earlier work this paper cites.
Neuroanimator: Fast neural network emulation and control of physics-based models
Radek Grzeszczuk, Demetri Terzopoulos, and Geoffrey Hinton · 1998
Earlier work this paper cites.
Computing the physical parameters of rigid-body motion from video
Kiran S Bhat, Steven M Seitz, Jovan Popović, and Pradeep K Khosla · 2002
Earlier work this paper cites.
Estimating contact dynamics
Marcus A Brubaker, Leonid Sigal, and David J Fleet · 2009
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini · 2009
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
Marc Deisenroth and Carl E Rasmussen · 2011
Earlier work this paper cites.
Internal physics models guide probabilistic judgments about object dynamics
Jessica Hamrick, Peter Battaglia, and Joshua B Tenenbaum · 2011
Earlier work this paper cites.
Nonlinear model predictive control
Frank Allgöwer and Alex Zheng · 2012
Earlier work this paper cites.
Model predictive control
Eduardo F Camacho and Carlos Bordons Alba · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Learning visual predictive models of physics for playing billiards
Katerina Fragkiadaki, Pulkit Agrawal, Sergey Levine, and Jitendra Malik · 2015
Earlier work this paper cites.
Fast R-CNN
Ross Girshick · 2015
Earlier work this paper cites.
3d reasoning from blocks to stability
Zhaoyin Jia, Andrew C Gallagher, Ashutosh Saxena, and Tsuhan Chen · 2015
Earlier work this paper cites.
Action-conditional video prediction using deep networks in atari games
Junhyuk Oh, Xiaoxiao Guo, Honglak Lee, Richard L Lewis, and Satinder Singh · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
From pixels to torques: Policy learning with deep dynamical models
Niklas Wahlström, Thomas B Schön, and Marc Peter Deisenroth · 2015
Earlier work this paper cites.
Galileo: Perceiving physical object properties by integrating a physics engine with deep learning
Jiajun Wu, Ilker Yildirim, Joseph J Lim, Bill Freeman, and Josh Tenenbaum · 2015
Earlier work this paper cites.
Learning to poke by poking: Experiential learning of intuitive physics
Pulkit Agrawal, Ashvin V Nair, Pieter Abbeel, Jitendra Malik, and Sergey Levine · 2016
Earlier work this paper cites.
Interaction networks for learning about objects, relations and physics
Peter Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, et al · 2016
Earlier work this paper cites.
Long-term image boundary extrapolation
Apratim Bhattacharyya, Mateusz Malinowski, Bernt Schiele, and Mario Fritz · 2016
Cited alongside, same era.
A compositional object-based approach to learning physical dynamics
Michael B Chang, Tomer Ullman, Antonio Torralba, and Joshua B Tenenbaum · 2016
Cited alongside, same era.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Cited alongside, same era.
Dynamic filter networks
Xu Jia, Bert De Brabandere, Tinne Tuytelaars, and Luc V Gool · 2016
Cited alongside, same era.
Learning physical intuition of block towers by example
Adam Lerer, Sam Gross, and Rob Fergus · 2016
Cited alongside, same era.
Sgdr: Stochastic gradient descent with warm restarts
Ilya Loshchilov and Frank Hutter · 2016
Cited alongside, same era.
Mocogan: Decomposing motion and content for video generation
Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang, and Jan Kautz · 2017
Later among the works it cites.
The pose knows: Video forecasting by generating pose futures
Jacob Walker, Kenneth Marino, Abhinav Gupta, and Martial Hebert · 2017
Later among the works it cites.
Visual interaction networks: Learning a physics simulator from video
Nicholas Watters, Daniel Zoran, Theophane Weber, Peter Battaglia, Razvan Pascanu, and Andrea Tacchetti · 2017
Later among the works it cites.
Learning to see physics via visual de-animation
Jiajun Wu, Erika Lu, Pushmeet Kohli, Bill Freeman, and Josh Tenenbaum · 2017
Later among the works it cites.
Relational inductive biases, deep learning, and graph networks
Peter W Battaglia, Jessica B Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez, Vinicius Zambaldi, Mateusz Malinowski, Andrea Tacchetti, David Raposo, Adam Santoro, Ryan Faulkner, et al · 2018
Later among the works it cites.
Stochastic video generation with a learned prior
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stacked hourglass networks for human pose estimation
Alejandro Newell, Kaiyu Yang, and Jia Deng · 2016
Cited alongside, same era.
Generating videos with scene dynamics
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba · 2016
Cited alongside, same era.
An uncertain future: Forecasting from static images using variational autoencoders
Jacob Walker, Carl Doersch, Abhinav Gupta, and Martial Hebert · 2016
Cited alongside, same era.
Physics 101: Learning physical object properties from unlabeled videos
Jiajun Wu, Joseph J Lim, Hongyi Zhang, Joshua B Tenenbaum, and William T Freeman · 2016
Cited alongside, same era.
Visual dynamics: Probabilistic future frame synthesis via cross convolutional networks
Tianfan Xue, Jiajun Wu, Katherine Bouman, and Bill Freeman · 2016
Cited alongside, same era.
Stochastic variational video prediction
Mohammad Babaeizadeh, Chelsea Finn, Dumitru Erhan, Roy H Campbell, and Sergey Levine · 2017
Cited alongside, same era.
Emily Denton and Rob Fergus · 2018
Later among the works it cites.
Shapestacks: Learning vision-based physical intuition for generalised object stacking
Oliver Groth, Fabian Fuchs, Ingmar Posner, and Andrea Vedaldi · 2018
Later among the works it cites.
Stochastic adversarial video prediction
Alex X Lee, Richard Zhang, Frederik Ebert, Pieter Abbeel, Chelsea Finn, and Sergey Levine · 2018
Later among the works it cites.
Zero-shot visual imitation
Deepak Pathak, Parsa Mahmoudieh, Guanghao Luo, Pulkit Agrawal, Dian Chen, Yide Shentu, Evan Shelhamer, Jitendra Malik, Alexei A Efros, and Trevor Darrell · 2018
Later among the works it cites.
Interpretable intuitive physics model
Tian Ye, Xiaolong Wang, James Davidson, and Abhinav Gupta · 2018
Later among the works it cites.
Phyre: A new benchmark for physical reasoning
Anton Bakhtin, Laurens van der Maaten, Justin Johnson, Laura Gustafson, and Ross Girshick · 2019
Later among the works it cites.
Large-scale study of curiosity-driven learning
Yuri Burda, Harri Edwards, Deepak Pathak, Amos Storkey, Trevor Darrell, and Alexei A. Efros · 2019
Later among the works it cites.
Mesh R-CNN
Georgia Gkioxari, Jitendra Malik, and Justin Johnson · 2019
Later among the works it cites.
Reasoning about physical interactions with object-oriented prediction and planning
Michael Janner, Sergey Levine, William T Freeman, Joshua B Tenenbaum, Chelsea Finn, and Jiajun Wu · 2019
Later among the works it cites.
Time-agnostic prediction: Predicting predictable video frames
Dinesh Jayaraman, Frederik Ebert, Alexei A Efros, and Sergey Levine · 2019
Later among the works it cites.
Propagation networks for model-based control under partial observation
Yunzhu Li, Jiajun Wu, Jun-Yan Zhu, Joshua B Tenenbaum, Antonio Torralba, and Russ Tedrake · 2019
Later among the works it cites.
Entity abstraction in visual model-based reinforcement learning
Rishi Veerapaneni, John D Co-Reyes, Michael Chang, Michael Janner, Chelsea Finn, Jiajun Wu, Joshua B Tenenbaum, and Sergey Levine · 2019
Later among the works it cites.
Compositional video prediction
Yufei Ye, Maneesh Singh, Abhinav Gupta, and Shubham Tulsiani · 2019
Later among the works it cites.
Forward prediction for physical reasoning
Rohit Girdhar, Laura Gustafson, Aaron Adcock, and Laurens van der Maaten · 2020
Closest in time.
Contrastive learning of structured world models
Thomas Kipf, Elise van der Pol, and Max Welling · 2020
Closest in time.
Learning to simulate complex physics with graph networks
Alvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff, Rex Ying, Jure Leskovec, and Peter W Battaglia · 2020
Closest in time.
Clevrer: Collision events for video representation and reasoning
Kexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli, Jiajun Wu, Antonio Torralba, and Joshua B Tenenbaum · 2020
Closest in time.