Fetching the paper…
Reading the bibliography…
This paper introduces MDP homomorphic networks for deep reinforcement learning.
Dynamic Programming
Richard E. Bellman · 1957
Earlier work this paper cites.
Neuronlike adaptive elements that can solve difficult learning control problems
Andrew G. Barto, Richard S. Sutton, and Charles W. Anderson · 1983
Earlier work this paper cites.
Temporal difference learning of position evaluation in the game of Go
Nicol N. Schraudolph, Peter Dayan, and Terrence J. Sejnowski · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Symmetries and model minimization in Markov Decision Processes
Balaraman Ravindran and Andrew G. Barto · 2001
Earlier work this paper cites.
SMDP homomorphisms: An algebraic approach to abstraction in Semi Markov Decision Processes
Balaraman Ravindran and Andrew G. Barto · 2003
Earlier work this paper cites.
Abstract Algebra
David Steven Dummit and Richard M. Foote · 2004
Earlier work this paper cites.
Approximate homomorphisms: A framework for non-exact minimization in Markov Decision Processes
Balaraman Ravindran and Andrew G. Barto · 2004
Earlier work this paper cites.
Towards a unified theory of state abstraction for mdps
Lihong Li, Thomas J. Walsh, and Michael L. Littman · 2006
Earlier work this paper cites.
On the hardness of finding symmetries in Markov decision processes
Shravan Matthur Narayanamurthy and Balaraman Ravindran · 2008
Earlier work this paper cites.
Bounding performance loss in approximate MDP homomorphisms
Jonathan Taylor, Doina Precup, and Prakash Panagaden · 2008
Earlier work this paper cites.
Teaching deep convolutional neural networks to play Go
Christopher Clark and Amos Storkey · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Autonomous weapons: An open letter from AI & robotics researchers, 2015
Future of Life Institute · 2015
Earlier work this paper cites.
Concrete problems in AI safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Group equivariant convolutional networks
Taco S. Cohen and Max Welling · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza, Alex Graves, Tim Harley, Timothy P. Lillicrap, David Silver, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2016
Cited alongside, same era.
Steerable CNNs
Taco S. Cohen and Max Welling · 2017
Cited alongside, same era.
Symmetry learning for function approximation in reinforcement learning
Online abstraction with MDP homomorphisms for deep learning
Ondrej Biza and Robert Platt · 2019
Later among the works it cites.
Quantifying generalization in reinforcement learning
Karl Cobbe, Oleg Klimov, Chris Hesse, Taehoon Kim, and John Schulman · 2019
Later among the works it cites.
A general theory of equivariant CNNs on homogeneous spaces
Taco S. Cohen, Mario Geiger, and Maurice Weiler · 2019
Later among the works it cites.
Learning to convolve: A generalized weight-tying approach
Nichita Diaconu and Daniel E. Worrall · 2019
Later among the works it cites.
PIC: Permutation invariant critic for multi-agent deep reinforcement learning
Iou-Jen Liu, Raymond A. Yeh, and Alexander G. Schwing · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Anuj Mahajan and Theja Tulabandhula · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Harmonic networks: Deep translation and rotation equivariance
Daniel E. Worrall, Stephan J. Garbin, Daniyar Turmukhambetov, and Gabriel J. Brostow · 2017
Cited alongside, same era.
Roto-translation covariant convolutional networks for medical image analysis
Erik J. Bekkers, Maxime W. Lafarge, Mitko Veta, Koen A.J. Eppenhof, Josien P.W. Pluim, and Remco Duits · 2018
Cited alongside, same era.
Linear model predictive safety certification for learning-based control
K. P. Wabersich and M. N. Zeilinger · 2018
Cited alongside, same era.
3D steerable CNNs: Learning rotationally equivariant features in volumetric data
Maurice Weiler, Mario Geiger, Max Welling, Wouter Boomsma, and Taco S. Cohen · 2018
Cited alongside, same era.
Learning steerable filters for rotation equivariant CNNs
Maurice Weiler, Fred A. Hamprecht, and Martin Storath · 2018
Cited alongside, same era.
Later among the works it cites.
rlpyt: A research code base for deep reinforcement learning in Pytorch
Adam Stooke and Pieter Abbeel · 2019
Later among the works it cites.
General E(2)-equivariant steerable CNNs
Maurice Weiler and Gabriele Cesa · 2019
Later among the works it cites.
Deep scale-spaces: Equivariance over scale
Daniel E. Worrall and Max Welling · 2019
Later among the works it cites.
Graph convolutional reinforcement learning
Jiechuan Jiang, Chen Dun, Tiejun Huang, and Zongqing Lu · 2020
Closest in time.
Image augmentation is all you need: Regularizing deep reinforcement learning from pixels
Ilya Kostrikov, Denis Yarats, and Rob Fergus · 2020
Closest in time.
Reinforcement learning with augmented data
Michael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto, Pieter Abbeel, and Aravind Srinivas · 2020
Closest in time.
Network randomization: A simple technique for generalization in deep reinforcement learning
Kimin Lee, Kibok Lee, Jinwoo Shin, and Honglak Lee · 2020
Closest in time.
Invariant transform experience replay: Data augmentation for deep reinforcement learning
Yijiong Lin, Jiancong Huang, Matthieu Zimmer, Yisheng Guan, Juan Rojas, and Paul Weng · 2020
Closest in time.
Goal-conditioned batch reinforcement learning for rotation invariant locomotion
Aditi Mavalankar · 2020
Closest in time.
Plannable approximations to MDP homomorphisms: Equivariance under actions
Elise van der Pol, Thomas Kipf, Frans A. Oliehoek, and Max Welling · 2020
Closest in time.