Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) has achieved significant success in various robot tasks: manipulation, navigation, etc.
Planning using a temporal world model
J. F. Allen and J. A. Koomen · 1983
Earlier work this paper cites.
A decision-theoretic approach to planning, perception, and control
K. Basye, T. Dean, J. Kirman, and M. Lejter · 1992
Earlier work this paper cites.
Multiple model-based reinforcement learning
K. Doya, K. Samejima, K.-i. Katagiri, and M. Kawato · 2002
Earlier work this paper cites.
A tutorial on energy-based learning
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang · 2006
Earlier work this paper cites.
Underactuated robotics: Learning, planning, and control for efficient and agile machines course notes for mit 6.832
R. Tedrake · 2009
Earlier work this paper cites.
A fast and simple algorithm for training neural probabilistic language models
A. Mnih and Y. W. Teh · 2012
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Model predictive control
E. F. Camacho and C. B. Alba · 2013
Earlier work this paper cites.
Local fisher discriminant analysis for pedestrian re-identification
S. Pedagadi, J. Orwell, S. Velastin, and B. Boghossian · 2013
Earlier work this paper cites.
Auto-encoding variational Bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
The mit super mini cheetah: A small, low-cost quadrupedal robot for dynamic locomotion
W. Bosworth, S. Kim, and N. Hogan · 2015
Earlier work this paper cites.
A recurrent latent variable model for sequential data
J. Chung, K. Kastner, L. Dinh, K. Goel, A. C. Courville, and Y. Bengio · 2015
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Learning to poke by poking: Experiential learning of intuitive physics
P. Agrawal, A. V. Nair, P. Abbeel, J. Malik, and S. Levine · 2016
Cited alongside, same era.
node2vec: Scalable feature learning for networks
A. Grover and J. Leskovec · 2016
Cited alongside, same era.
Deep variational information bottleneck
A. A. Alemi, I. Fischer, J. V. Dillon, and K. Murphy · 2016
Cited alongside, same era.
Mastering the game of go without human knowledge
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al · 2017
Cited alongside, same era.
QMDP-net: Deep learning for planning under partial observability
Particle filter networks with application to visual localization
P. Karkus, D. Hsu, and W. S. Lee · 2018
Later among the works it cites.
Representation learning with contrastive predictive coding
A. v. d. Oord, Y. Li, and O. Vinyals · 2018
Later among the works it cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Distributed distributional deterministic policy gradients
G. Barth-Maron, M. W. Hoffman, D. Budden, W. Dabney, D. Horgan, D. Tb, A. Muldal, N. Heess, and T. Lillicrap · 2018
Later among the works it cites.
Pybullet, a python module for physics simulation for games, robotics and machine learning
E. Coumans and Y. Bai · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Karkus, D. Hsu, and W. S. Lee · 2017
Cited alongside, same era.
Deep visual foresight for planning robot motion
C. Finn and S. Levine · 2017
Cited alongside, same era.
Self-supervised deep reinforcement learning with generalized computation graphs for robot navigation
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Recurrent world models facilitate policy evolution
D. Ha and J. Schmidhuber · 2018
Cited alongside, same era.
Deep variational reinforcement learning for POMDPs
M. Igl, L. Zintgraf, T. A. Le, F. Wood, and S. Whiteson · 2018
Cited alongside, same era.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2018
Cited alongside, same era.
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, D. d. L. Casas, D. Budden, A. Abdolmaleki, J. Merel, A. Lefrancq, et al · 2018
Cited alongside, same era.
Later among the works it cites.
Dream to control: Learning behaviors by latent imagination
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi · 2019
Later among the works it cites.
Deep graph infomax
P. Velickovic, W. Fedus, W. L. Hamilton, P. Liò, Y. Bengio, and R. D. Hjelm · 2019
Later among the works it cites.
Differentiable algorithm networks for composable robot learning
P. Karkus, X. Ma, D. Hsu, L. P. Kaelbling, W. S. Lee, and T. Lozano-Pérez · 2019
Later among the works it cites.
Particle filter recurrent neural networks
X. Ma, P. Karkus, D. Hsu, and W. S. Lee · 2020
Closest in time.
Contrastive learning of structured world models
T. Kipf, E. van der Pol, and M. Welling · 2020
Closest in time.
Discriminative particle filter reinforcement learning for complex partial observations
X. Ma, P. Karkus, D. Hsu, W. S. Lee, and N. Ye · 2020
Closest in time.
P. Karkus, A. Angelova, V. Vanhoucke, and R. Jonschkowski · 2020
Closest in time.