Fetching the paper…
Reading the bibliography…
Humans easily recognize object parts and their hierarchical structure by watching how they move; they can then predict how each part moves in the future.
Visual perception of biological motion and a model for its analysis
Gunnar Johansson · 1973
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Geoffrey E Hinton and Drew Van Camp · 1993
Earlier work this paper cites.
Layered representation for motion analysis
John YA Wang and Edward H Adelson · 1993
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Learning parts-based representations of data
David A Ross and Richard S Zemel · 2006
Earlier work this paper cites.
Core knowledge
Elizabeth S Spelke and Katherine D Kinzler · 2007
Earlier work this paper cites.
Physics-based person tracking using the anthropomorphic walker
Marcus A. Brubaker, David J. Fleet, and Aaron Hertzmann · 2009
Earlier work this paper cites.
The origin of concepts
Susan Carey · 2009
Earlier work this paper cites.
Beyond pixels: exploring new representations and applications for motion analysis
Ce Liu · 2009
Earlier work this paper cites.
Efficient hierarchical graph-based video segmentation
Matthias Grundmann, Vivek Kwatra, Mei Han, and Irfan Essa · 2010
Earlier work this paper cites.
Learning articulated structure and motion
David A Ross, Daniel Tarlow, and Richard S Zemel · 2010
Earlier work this paper cites.
Physically-based motion models for 3d tracking: A convex formulation
Mathieu Salzmann and Raquel Urtasun · 2011
Earlier work this paper cites.
Streaming hierarchical video segmentation
Chenliang Xu, Caiming Xiong, and Jason J Corso · 2012
Earlier work this paper cites.
Fast rigid motion segmentation via incrementally-complex local models
Fernando Flores-Mangas and Allan D Jepson · 2013
Earlier work this paper cites.
Principles of Gestalt psychology
Kurt Koffka · 2013
Earlier work this paper cites.
Physically plausible 3d scene tracking: The single actor hypothesis
Nikolaos Kyriazis and Antonis Argyros · 2013
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Earlier work this paper cites.
Dynamical simulation priors for human motion tracking
Marek Vondrak, Leonid Sigal, and Odest Chadwicke Jenkins · 2013
Earlier work this paper cites.
Action localization with tubelets from motion
Mihir Jain, Jan Van Gemert, Hervé Jégou, Patrick Bouthemy, and Cees GM Snoek · 2014
Earlier work this paper cites.
Segmentation of moving objects by long term video analysis
Peter Ochs, Jitendra Malik, and Thomas Brox · 2014
Cited alongside, same era.
Imagining the unseen: Stability-based cuboid arrangements for scene understanding
Tianjia Shao, Aron Monszpart, Youyi Zheng, Bongjin Koo, Weiwei Xu, Kun Zhou, and Niloy J Mitra · 2014
Cited alongside, same era.
Multiple object recognition with visual attention
Jimmy Ba, Volodymyr Mnih, and Koray Kavukcuoglu · 2015
Cited alongside, same era.
Efficient inference in occlusion-aware generative models of images
Jonathan Huang and Kevin Murphy · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
A compositional object-based approach to learning physical dynamics
Michael B Chang, Tomer Ullman, Antonio Torralba, and Joshua B Tenenbaum · 2017
Later among the works it cites.
Xception: Deep learning with depthwise separable convolutions
François Chollet · 2017
Later among the works it cites.
Learning a physical long-term predictor
Sebastien Ehrhardt, Aron Monszpart, Niloy J Mitra, and Andrea Vedaldi · 2017
Later among the works it cites.
Neural expectation maximization
Klaus Greff, Sjoerd van Steenkiste, and Jürgen Schmidhuber · 2017
Later among the works it cites.
Beta-vae: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, Christopher P Burgess, Xavier Glorot, Matthew Botvinick, Shakir Mohamed, and Alexander Lerchner · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Cited alongside, same era.
Learning to poke by poking: Experiential learning of intuitive physics
Pulkit Agrawal, Ashvin Nair, Pieter Abbeel, Jitendra Malik, and Sergey Levine · 2016
Cited alongside, same era.
Interaction networks for learning about objects, relations and physics
Peter W. Battaglia, Razvan Pascanu, Matthew Lai, Danilo Rezende, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel · 2016
Cited alongside, same era.
Attend, infer, repeat: Fast scene understanding with generative models
SM Eslami, Nicolas Heess, Theophane Weber, Yuval Tassa, Koray Kavukcuoglu, and Geoffrey E Hinton · 2016
Cited alongside, same era.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Cited alongside, same era.
Wei-Ning Hsu, Yu Zhang, and James Glass · 2017
Later among the works it cites.
Activity forecasting
Kris M. Kitani, De-An Huang, and Wei-Chiu Ma · 2017
Later among the works it cites.
Deep predictive coding networks for video prediction and unsupervised learning
William Lotter, Gabriel Kreiman, and David Cox · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Later among the works it cites.
Visual interaction networks
Nicholas Watters, Andrea Tacchetti, Theophane Weber, Razvan Pascanu, Peter Battaglia, and Daniel Zoran · 2017
Later among the works it cites.
Learning to see physics via visual de-animation
Jiajun Wu, Erika Lu, Pushmeet Kohli, William T Freeman, and Joshua B Tenenbaum · 2017
Later among the works it cites.
Isolating sources of disentanglement in variational autoencoders
Tian Qi Chen, Xuechen Li, Roger Grosse, and David Duvenaud · 2018
Later among the works it cites.
Hierarchical disentangled representations
Babak Esmaeili, Hao Wu, Sarthak Jain, Siddharth Narayanaswamy, Brooks Paige, and Jan-Willem van de Meent · 2018
Later among the works it cites.
Scan: learning abstract hierarchical compositional visual concepts
Irina Higgins, Nicolas Sonnerat, Loic Matthey, Arka Pal, Christopher P Burgess, Matthew Botvinick, Demis Hassabis, and Alexander Lerchner · 2018
Later among the works it cites.
Neural relational inference for interacting systems
Thomas N Kipf, Ethan Fetaya, Kuan-Chieh Wang, Max Welling, and Richard S Zemel · 2018
Later among the works it cites.
Msnet: Mutual suppression network for disentangled video representations
Jungbeom Lee, Jangho Lee, Sungmin Lee, and Sungroh Yoon · 2018
Later among the works it cites.
Flow-grounded spatial-temporal video prediction from still images
Yijun Li, Chen Fang, Jimei Yang, Zhaowen Wang, Xin Lu, and Ming-Hsuan Yang · 2018
Later among the works it cites.
A deep generative model for disentangled representations of sequential data
Yingzhen Li and Stephan Mandt · 2018
Later among the works it cites.
Relational neural expectation maximization: Unsupervised discovery of objects and their interactions
Sjoerd van Steenkiste, Michael Chang, Klaus Greff, and Jürgen Schmidhuber · 2018
Later among the works it cites.