Fetching the paper…
Reading the bibliography…
Many Reinforcement Learning (RL) approaches use joint control signals (positions, velocities, torques) as action space for continuous control tasks.
J. J. Kuffner and S. M. LaValle, “Rrt-connect: An efficient approach to single-query path planning,” in Proceedings 2000 ICRA. Millennium Conference. IEEE International Conference on Robotics and Automation. Symposia Proceedings (Cat. No. 00CH37065) , vol. 2. IEEE, 2000, pp. 995–1001
2000
Earlier work this paper cites.
R. Bohlin and L. E. Kavraki, “Path planning using lazy prm,” in Proceedings 2000 ICRA. Millennium Conference. IEEE International Conference on Robotics and Automation. Symposia Proceedings (Cat. No. 00CH37065) , vol. 1. IEEE, 2000, pp. 521–528
2000
Earlier work this paper cites.
S. M. LaValle, Planning algorithms . Cambridge university press, 2006
2006
Earlier work this paper cites.
B. Siciliano and O. Khatib, Springer Handbook of Robotics . Berlin, Heidelberg: Springer-Verlag, 2007
2007
Earlier work this paper cites.
D. Berenson, J. Kuffner, and H. Choset, “An optimization approach to planning for mobile manipulation,” in 2008 IEEE International Conference on Robotics and Automation . IEEE, 2008, pp. 1187–1192
2008
Earlier work this paper cites.
N. Jetchev and M. Toussaint, “Trajectory prediction in cluttered voxel environments,” in 2010 IEEE International Conference on Robotics and Automation . IEEE, 2010, pp. 2523–2528
2010
Earlier work this paper cites.
E. Klingbeil, A. Saxena, and A. Y. Ng, “Learning to open new doors,” in 2010 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2010, pp. 2751–2757
2010
Earlier work this paper cites.
A. Dragan, G. J. Gordon, and S. Srinivasa, “Learning from experience in manipulation planning: Setting the right goals,” in In Proceedings of the ISRR , 2011
2011
Earlier work this paper cites.
G. Konidaris, S. Kuindersma, R. Grupen, and A. Barto, “Autonomous skill acquisition on a mobile manipulator,” in Twenty-Fifth AAAI Conference on Artificial Intelligence , 2011
2011
Earlier work this paper cites.
2013
Earlier work this paper cites.
S. Levine and V. Koltun, “Guided policy search,” in International Conference on Machine Learning , 2013, pp. 1–9
2013
Earlier work this paper cites.
2015
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , vol. 17, no. 1, pp. 1334–1373, 2016
2016
Earlier work this paper cites.
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy, “Deep exploration via bootstrapped dqn,” in Advances in neural information processing systems , 2016, pp. 4026–4034
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
T. D. Kulkarni, K. Narasimhan, A. Saeedi, and J. Tenenbaum, “Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation,” in Advances in neural information processing systems , 2016, pp. 3675–3683
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Earlier work this paper cites.
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi, “Target-driven visual navigation in indoor scenes using deep reinforcement learning,” in 2017 IEEE international conference on robotics and automation (ICRA) . IEEE, 2017, pp. 3357–3364
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. S. Vezhnevets, S. Osindero, T. Schaul, N. Heess, M. Jaderberg, D. Silver, and K. Kavukcuoglu, “Feudal networks for hierarchical reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 2017, pp. 3540–3549
2017
Earlier work this paper cites.
D. Hernandez, “How to survive a robot apocalypse: Just close the door,” The Wall Street Journal , p. 10, 2017
2017
Cited alongside, same era.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Cited alongside, same era.
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke et al. , “Scalable deep reinforcement learning for vision-based robotic manipulation,” in Conference on Robot Learning , 2018, pp. 651–673
2018
Cited alongside, same era.
A. Zeng, S. Song, S. Welker, J. Lee, A. Rodriguez, and T. Funkhouser, “Learning synergies between pushing and grasping with self-supervised deep reinforcement learning,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 4238–4245
2018
Cited alongside, same era.
A. H. Qureshi, A. Simeonov, M. J. Bency, and M. C. Yip, “Motion planning networks,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 2118–2124
2019
Later among the works it cites.
S. Bansal, V. Tolani, S. Gupta, J. Malik, and C. Tomlin, “Combining optimal control and learning for visual navigation in novel environments,” in Conference on Robot Learning (CoRL) , 2019
2019
Later among the works it cites.
Y. Jiang, F. Yang, S. Zhang, and P. Stone, “Integrating task-motion planning with reinforcement learning for robust decision making in mobile robots,” in In Proceedings of the AAMAS , 2019
2019
Later among the works it cites.
K. Ciosek, Q. Vuong, R. Loftin, and K. Hofmann, “Better exploration with optimistic actor critic,” in Advances in Neural Information Processing Systems , 2019, pp. 1787–1798
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
O. Nachum, S. S. Gu, H. Lee, and S. Levine, “Data-efficient hierarchical reinforcement learning,” in Advances in Neural Information Processing Systems , 2018, pp. 3303–3313
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
M. Rana, M. Mukadam, S. R. Ahmadzadeh, S. Chernova, and B. Boots, “Towards robust skill generalization: Unifying learning from demonstration and motion planning,” in Intelligent robots and systems , 2018
2018
Cited alongside, same era.
O. Nachum, S. Gu, H. Lee, and S. Levine, “Near-optimal representation learning for hierarchical reinforcement learning,” International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
I. Osband, J. Aslanides, and A. Cassirer, “Randomized prior functions for deep reinforcement learning,” in Advances in Neural Information Processing Systems , 2018, pp. 8617–8629
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Sergio Guadarrama and others, “TF-Agents: A library for reinforcement learning in tensorflow,” https://github.com/tensorflow/agents , 2018. [Online]. Available: https://github.com/tensorflow/agents
2018
Cited alongside, same era.
Manolis Savva, Abhishek Kadian, Oleksandr Maksymets et al. , “Habitat: A Platform for Embodied AI Research,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
X. Wang, B. Zhou, Y. Shi, X. Chen, Q. Zhao, and K. Xu, “Shape2motion: Joint analysis of motion parts and attributes from 3d shapes,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 8876–8884
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
Y. Chebotar, A. Handa, V. Makoviychuk, M. Macklin, J. Issac, N. Ratliff, and D. Fox, “Closing the sim-to-real loop: Adapting simulation randomization with real world experience,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 8973–8979
2019
Later among the works it cites.
K. Kang, S. Belkhale, G. Kahn, P. Abbeel, and S. Levine, “Generalization through simulation: Integrating simulated and real data into deep reinforcement learning for vision-based autonomous flight,” International Conference on Robotics and Automation (ICRA) , 2019
2019
Later among the works it cites.
X. Meng, N. Ratliff, Y. Xiang, and D. Fox, “Neural autonomous navigation with riemannian motion policy,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 8860–8866
2019
Later among the works it cites.
H. Quan, Y. Li, and Y. Zhang, “A novel mobile robot navigation method based on deep reinforcement learning,” International Journal of Advanced Robotic Systems , vol. 17, no. 3, p. 1729881420921672, 2020
2020
Closest in time.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,” IEEE Transactions on Robotics , 2020
2020
Closest in time.
C. Li, F. Xia, R. Martín-Martín, and S. Savarese, “Hrl4in: Hierarchical reinforcement learning for interactive navigation with mobile manipulators,” in Conference on Robot Learning , 2020, pp. 603–616
2020
Closest in time.
2020
Closest in time.
J. Yamada, G. Salhotra, Y. Lee, M. Pflueger, K. Pertsch, P. Englert, G. S. Sukhatme, and J. J. Lim, “Motion planner augmented action spaces for reinforcement learning,” RSS Workshop on Action Representations for Learning in Continuous Control , 2020
2020
Closest in time.
J. Wu, X. Sun, A. Zeng, S. Song, J. Lee, S. Rusinkiewicz, and T. Funkhouser, “Spatial Action Maps for Mobile Manipulation,” in Proceedings of Robotics: Science and Systems , Corvalis, Oregon, USA, July 2020
2020
Closest in time.
F. Xia, W. B. Shen, C. Li, P. Kasimbeg, M. E. Tchapmi, A. Toshev, R. Martín-Martín, and S. Savarese, “Interactive gibson benchmark: A benchmark for interactive navigation in cluttered environments,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 713–720, April 2020
2020
Closest in time.
C. Wang, Q. Zhang, Q. Tian, S. Li, X. Wang, D. Lane, Y. Petillot, and S. Wang, “Learning mobile manipulation through deep reinforcement learning,” Sensors , vol. 20, no. 3, p. 939, 2020
2020
Closest in time.
K. Rao, C. Harris, A. Irpan, S. Levine, J. Ibarz, and M. Khansari, “Rl-cyclegan: Reinforcement learning aware simulation-to-real,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 11 157–11 166
2020
Closest in time.