Fetching the paper…
Reading the bibliography…
This work exploits action equivariance for representation learning in reinforcement learning.
Nicholas Watters, Loic Matthey, Matko Bosnjak, Christopher P Burgess, and Alexander Lerchner. 2019 · 1905
Earlier work this paper cites.
Algebraic Structure Theory Of Sequential Machines
Juris Hartmanis and R. E. Stearns. 1966 · 1966
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Martin L Puterman. 1994 · 1994
Earlier work this paper cites.
Model Minimization in Markov Decision Processes. In AAAI Conference on Artifical Intelligence/IAAI Conference on Innovative Applications of Artificial Intelligence
Thomas Dean and Robert Givan. 1997 · 1997
Earlier work this paper cites.
Symmetries and Model Minimization in Markov Decision Processes
Balaraman Ravindran and Andrew G. Barto. 2001 · 2001
Earlier work this paper cites.
Equivalence Notions and Model Minimization in Markov Decision Processes. In Artificial Intelligence
Robert Givan, Thomas Dean, and Matthew Greig. 2003 · 2003
Earlier work this paper cites.
Metrics for Finite Markov Decision Processes. In Conference on Uncertainty in Artificial Intelligence
Norm Ferns, Prakash Panangaden, and Doina Precup. 2004 · 2004
Earlier work this paper cites.
Approximate Homomorphisms: A Framework for Non-Exact Minimization in Markov Decision Processes. In International Conference on Knowledge Based Computer Systems
Balaraman Ravindran and Andrew G. Barto. 2004 · 2004
Earlier work this paper cites.
Towards a Unified Theory of State Abstraction for MDPs. In International Symposium on Artificial Intelligence and Mathematics
Lihong Li, Thomas J. Walsh, and Michael L. Littman. 2006 · 2006
Earlier work this paper cites.
Bounding Performance Loss in Approximate MDP Homomorphisms. In Advances in Neural Information Processing Systems
Jonathan Taylor, Doina Precup, and Prakash Panagaden. 2008 · 2008
Earlier work this paper cites.
Auto-Encoding Variational Bayes. In International Conference on Learning Representations
Diederik P. Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Learning State Representations with Robotic Priors. In Autonomous Robots
Rico Jonschkowski and Oliver Brock. 2015 · 2015
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization. In International Conference on Learning Representations
Diederik P. Kingma and Jimmy Lei Ba. 2015 · 2015
Earlier work this paper cites.
Human-level Control through Deep Reinforcement Learning. In Nature
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis. 2015 · 2015
Earlier work this paper cites.
Embed to Control: a Locally Linear Latent Dynamics Model for Control from Raw Images. In Advances in Neural Information Processing Systems
Manuel Watter, Jost Tobias Springenberg, Joschka Boedecker, and Martin Riedmiller. 2015 · 2015
Earlier work this paper cites.
Learning to Poke by Poking: Experiential Learning of Intuitive Physics. In Advances in Neural Information Processing Systems
Pulkit Agrawal, Ashvin Nair, Pieter Abbeel, Jitendra Malik, and Sergey Levine. 2016 · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Earlier work this paper cites.
Group Equivariant Convolutional Networks. In International Conference on Machine Learning
Taco S. Cohen and Max Welling. 2016 · 2016
Earlier work this paper cites.
Asynchronous Methods for Deep Reinforcement Learning. In International Conference on Machine Learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza, Alex Graves, Tim Harley, Timothy P. Lillicrap, David Silver, and Koray Kavukcuoglu. 2016 · 2016
Cited alongside, same era.
Value Iteration Networks. In Advances in Neural Information Processing Systems
Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Reinforcement Learning with Unsupervised Auxiliary Tasks. In International Conference on Learning Representations
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z. Leibo, David Silver, and Koray Kavukcuoglu. 2017 · 2017
Cited alongside, same era.
Categorical Reparameterization with Gumbel-Softmax. In International Conference on Learning Representations
Eric Jang, Shixiang Gu, and Ben Poole. 2017 · 2017
Cited alongside, same era.
QMDP-net: Deep Learning for Planning under Partial Observability. In Advances in Neural Information Processing Systems
Peter Karkus, David Hsu, and Wee Sun Lee. 2017 · 2017
Neural Relational Inference for Interacting Systems. In International Conference on Machine Learning
Thomas Kipf, Ethan Fetaya, Kuan-Chieh Wang, Max Welling, and Richard Zemel. 2018 · 2018
Later among the works it cites.
Diffusion-Based Approximate Value Functions. In ICML ECA Workshop
Martin Klissarov and Doina Precup. 2018 · 2018
Later among the works it cites.
Gated Path Planning Networks. In International Conference on Machine Learning
Lisa Lee, Emilio Parisotto, Devendra Singh Chaplot, Eric P. Xing, and Ruslan Salakhutdinov. 2018 · 2018
Later among the works it cites.
Generalized Value Iteration Networks: Life Beyond Lattices. In AAAI Conference on Artificial Intelligence
Sufeng Niu, Siheng Chen, Colin Targonski, Melissa Smith, Jelena Kova evi, and Hanyu Guo. 2018 · 2018
Later among the works it cites.
Representation Learning with Contrastive Predictive Coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Value Prediction Network. In Advances in Neural Information Processing Systems
Junhyuk Oh, Satinder Singh, and Honglak Lee. 2017 · 2017
Cited alongside, same era.
Independently Controllable Factors
Valentin Thomas, Jules Pondard, Emmanuel Bengio, Marc Sarfati, Philippe Beaudoin, Marie-Jean Meurs, Joelle Pineau, Doina Precup, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Harmonic Networks: Deep Translation and Rotation Equivariance. In IEEE Conference on Computer Vision and Pattern Recognition
Daniel E. Worrall, Stephan J. Garbin, Daniyar Turmukhambetov, and Gabriel J. Brostow. 2017 · 2017
Cited alongside, same era.
Fashion-MNIST: A Novel Image Dataset for Benchmarking Machine Learning Algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf. 2017 · 2017
Cited alongside, same era.
Playing Hard Exploration Games by Watching YouTube. In Advances in Neural Information Processing Systems
Yusuf Aytar, Tobias Pfaff, David Budden, Thomas Paine, Ziyu Wang, and Nando de Freitas. 2018 · 2018
Cited alongside, same era.
Efficient Model-Based Deep Reinforcement Learning with Variational State Tabulation. In International Conference on Machine Learning
Dane Corneil, Wulfram Gerstner, and Johanni Brea. 2018 · 2018
Cited alongside, same era.
TreeQN and ATreeC: Differentiable Tree-Structured Models for Deep Reinforcement Learning. In International Conference on Learning Representations
Gregory Farquhar, Tim Rocktäschel, Maximilian Igl, and SA Whiteson. 2018 · 2018
Cited alongside, same era.
Composable Planning with Attributes. In International Conference on Machine Learning
Amy Zhang, Adam Lerer, Sainbayar Sukhbaatar, Rob Fergus, and Arthur Szlam. 2018 · 2018
Later among the works it cites.
Unsupervised State Representation Learning in Atari. In Advances in Neural Information Processing Systems
Ankesh Anand, Evan Racah, Sherjil Ozair, Yoshua Bengio, Marc-Alexandre Cote, and R Devon Hjelm. 2019 · 2019
Later among the works it cites.
Unsupervised Grounding of Plannable First-Order Logic Representation from Images. In International Conference on Automated Planning and Scheduling
Masataro Asai. 2019 · 2019
Later among the works it cites.
Combined Reinforcement Learning via Abstract Representations. In AAAI Conference on Artificial Intelligence
Vincent François-Lavet, Yoshua Bengio, Doina Precup, and Joelle Pineau. 2019 · 2019
Later among the works it cites.
DeepMDP: Learning Continuous Latent Space Models for Representation Learning. In International Conference on Machine Learning
Carles Gelada, Saurabh Kumar, Jacob Buckman, Ofir Nachum, and Marc G. Bellemare. 2019 · 2019
Later among the works it cites.
Learning Actionable Representations with Goal Conditioned Policies. In International Conference on Learning Representations
Dibya Ghosh, Abhishek Gupta, and Sergey Levine. 2019 · 2019
Later among the works it cites.
Learning Latent Dynamics for Planning from Pixels. In International Conference on Machine Learning
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson. 2019 · 2019
Later among the works it cites.
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model
Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert, Karen Simonyan, Laurent Sifre, Simon Schmitt, Arthur Guez, Edward Lockhart, Demis Hassabis, Thore Graepel, Timothy P. Lillicrap, and David Silver. 2019 · 2019
Later among the works it cites.
Learning Robotic Manipulation through Visual Planning and Acting. In Robotics: Science and Systems
Angelina Wang, Thanard Kurutach, Kara Liu, Pieter Abbeel, and Aviv Tamar. 2019 · 2019
Later among the works it cites.
Marvin Zhang, Sharad Vikram, Laura Smith, Pieter Abbeel, Matthew Johnson, and Sergey Levine. 2019 · 2019
Later among the works it cites.
Model-Based Reinforcement Learning for Atari. In International Conference on Learning Representations
Lukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski, Roy H. Campbell, Konrad Czechowski, Dumitru Erhan, Chelsea Finn, Piotr Kozakowski, Sergey Levine, et al · 2020
Closest in time.
Contrastive Learning of Structured World Models. In International Conference on Learning Representations
Thomas Kipf, Elise van der Pol, and Max Welling. 2020 · 2020
Closest in time.