Fetching the paper…
Reading the bibliography…
In order to practically implement the door opening task, a policy ought to be robust to a wide distribution of door types and environment settings.
Quasi-direct drive for low-cost compliant robotic manipulation
D. V. Gealy, S. McKinley, B. Yi, P. Wu, P. R. Downey, G. Balke, A. Zhao, M. Guo, R. Thomasson, A. Sinclair, P. Cuellar, Z. McCarthy, and P. Abbeel · 1904
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Adam: A method for stochastic optimization, 2014
D. P. Kingma and J. Ba · 1992
Earlier work this paper cites.
Detecting and modeling doors with mobile robots
D. Anguelov, D. Koller, E. Parker, and S. Thrun · 2004
Earlier work this paper cites.
Shadowrobot dexterous hand, 2005
ShadowRobot · 2005
Earlier work this paper cites.
Laser-based perception for door and handle identification
R. B. Rusu, W. Meeussen, S. Chitta, and M. Beetz · 2009
Earlier work this paper cites.
Learning to open new doors
E. Klingbeil, A. Saxena, and A. Y. Ng · 2010
Earlier work this paper cites.
A generalized path integral control approach to reinforcement learning
E. Theodorou, J. Buchli, and S. Schaal · 2010
Earlier work this paper cites.
Learning force control policies for compliant manipulation
M. Kalakrishnan, L. Righetti, P. Pastor, and S. Schaal · 2011
Earlier work this paper cites.
“open sesame!” adaptive force/velocity control for opening unknown doors
Y. Karayiannidis, C. Smith, F. E. Viña, P. Ogren, and D. Kragic · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Learning the dynamics of doors for robotic manipulation
F. Endres, J. Trinkle, and W. Burgard · 2013
Earlier work this paper cites.
Autodesk fusion360, 2013
AutoDesk, Inc · 2013
Cited alongside, same era.
What happened at the darpa robotics challenge , and why ?
R. Du, S. Feng, P. Franklin, M. Gennert, J. P. Graff, P. He, A. Jaeger, J. Kim, K. Knoedler, L. Li, C. Y. Liu, X. Long, T. Padir, F. Polido, G. G. Tighe, and X. Xinjilefu · 2015
Cited alongside, same era.
Learning contact-rich manipulation skills with guided policy search
S. Levine, N. Wagener, and P. Abbeel · 2015
Cited alongside, same era.
Epopt: Learning robust neural network policies using model ensembles
A. Rajeswaran, S. Ghotra, S. Levine, and B. Ravindran · 2016
Cited alongside, same era.
(cad)$ˆ2$rl: Real single-image flight without a single real image
F. Sadeghi and S. Levine · 2016
Cited alongside, same era.
Sim2real view invariant visual servoing by recurrent control
F. Sadeghi, A. Toshev, E. Jang, and S. Levine · 2017
Later among the works it cites.
Sim-to-real transfer of robotic control with dynamics randomization
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel · 2017
Later among the works it cites.
Proximal policy optimization algorithms
P. D. A. R. O. K. John Schulman, Filip Wolski · 2017
Later among the works it cites.
Learning dexterous in-hand manipulation
OpenAI, M. Andrychowicz, B. Baker, M. Chociej, R. Józefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Openai gym, 2016
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
High-dimensional continuous control using generalized advantage estimation
J. Schulman, P. Moritz, S. Levine, M. Jordan, and P. Abbeel · 2016
Cited alongside, same era.
Door opening by joining reinforcement learning and intelligent control
B. Nemec, L. Žlajpah, and A. Ude · 2017
Cited alongside, same era.
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
S. Gu, E. Holly, T. Lillicrap, and S. Levine · 2017
Cited alongside, same era.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
A. Rajeswaran, V. Kumar, A. Gupta, J. Schulman, E. Todorov, and S. Levine · 2017
Cited alongside, same era.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Cited alongside, same era.
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, D. de Las Casas, D. Budden, A. Abdolmaleki, J. Merel, A. Lefrancq, T. P. Lillicrap, and M. A. Riedmiller · 2018
Later among the works it cites.
Roboturk: A crowdsourcing platform for robotic skill learning through imitation
A. Mandlekar, Y. Zhu, A. Garg, J. Booher, M. Spero, A. Tung, J. Gao, J. Emmons, A. Gupta, E. Orbay, S. Savarese, and L. Fei-Fei · 2018
Later among the works it cites.
Surreal: Open-source reinforcement learning framework and robot manipulation benchmark
L. Fan, Y. Zhu, J. Zhu, Z. Liu, O. Zeng, A. Gupta, J. Creus-Costa, S. Savarese, and L. Fei-Fei · 2018
Later among the works it cites.
Mujoco-py
OpenAI · 2018
Later among the works it cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Pytorch implementations of reinforcement learning algorithms
I. Kostrikov · 2018
Later among the works it cites.
Soft actor-critic algorithms and applications
T. Haarnoja, A. Zhou, K. Hartikainen, G. Tucker, S. Ha, J. Tan, V. Kumar, H. Zhu, A. Gupta, P. Abbeel, and S. Levine · 2018
Later among the works it cites.