Fetching the paper…
Reading the bibliography…
We present a novel approach to improve the performance of deep reinforcement learning (DRL) based outdoor robot navigation systems.
A note on two problems in connexion with graphs
E. W. Dijkstra et al · 1959
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Probabilistic roadmaps for path planning in high-dimensional configuration spaces
L. E. Kavraki, P. Svestka, J.-C. Latombe, and M. H. Overmars · 1996
Earlier work this paper cites.
The dynamic window approach to collision avoidance
D. Fox, W. Burgard, and S. Thrun · 1997
Earlier work this paper cites.
Path planning for autonomous vehicles driving over rough terrain
A. Lacaze, Y. Moscovitz, N. DeClaris, and K. Murphy · 1998
Earlier work this paper cites.
Rrt-connect: An efficient approach to single-query path planning
J. J. Kuffner and S. M. LaValle · 2000
Earlier work this paper cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz · 2015
Earlier work this paper cites.
Multirobot cooperative learning for semiautonomous control in urban search and rescue applications
Y. Liu and G. Nejat · 2016
Earlier work this paper cites.
Vime: Variational information maximizing exploration
R. Houthooft, X. Chen, Y. Duan, J. Schulman, F. De Turck, and P. Abbeel · 2016
Earlier work this paper cites.
Decentralized non-communicating multiagent collision avoidance with deep reinforcement learning
Y. F. Chen, M. Liu, M. Everett, and J. P. How · 2017
Earlier work this paper cites.
Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards
M. Vecerik, T. Hester, J. Scholz, F. Wang, O. Pietquin, B. Piot, N. Heess, T. Rothörl, T. Lampe, and M. Riedmiller · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
Deep reward shaping from demonstrations
A. Hussein, E. Elyan, M. M. Gaber, and C. Jayne · 2017
Earlier work this paper cites.
The beta policy for continuous control reinforcement learning
P.-W. Chou · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Earlier work this paper cites.
A critical investigation of deep reinforcement learning for navigation
V. Dhiman, S. Banerjee, B. Griffin, J. M. Siskind, and J. J. Corso · 2018
Earlier work this paper cites.
Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning
P. Long, T. Fan, X. Liao, W. Liu, H. Zhang, and J. Pan · 2018
Cited alongside, same era.
Robot navigation of environments with unknown rough terrain using deep reinforcement learning
K. Zhang, F. Niroui, M. Ficocelli, and G. Nejat · 2018
Cited alongside, same era.
Overcoming exploration in reinforcement learning with demonstrations
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel · 2018
Cited alongside, same era.
Deep reinforcement learning-based automatic exploration for navigation in unknown environment
H. Li, Q. Zhang, and D. Zhao · 2019
Cited alongside, same era.
Hierarchical automatic curriculum learning: Converting a sparse reward navigation task into dense reward
N. Jiang, S. Jin, and C. Zhang · 2019
Cited alongside, same era.
Planetary surface mobility and exploration: A review
A. Thoesen and H. Marvi · 2021
Later among the works it cites.
Deep reinforcement learning based mobile robot navigation: A review
K. Zhu and T. Zhang · 2021
Later among the works it cites.
Voila: Visual-observation-only imitation learning for autonomous navigation
H. Karnan, G. Warnell, X. Xiao, and P. Stone · 2021
Later among the works it cites.
Learning goal conditioned socially compliant navigation from demonstration using risk-based features
A. Konar, B. H. Baghi, and G. Dudek · 2021
Later among the works it cites.
On the sample complexity and metastability of heavy-tailed policy search in continuous control
A. S. Bedi, A. Parayil, J. Zhang, M. Wang, and A. Koppel · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H.-T. L. Chiang, A. Faust, M. Fiser, and A. Francis · 2019
Cited alongside, same era.
Dealing with sparse rewards in reinforcement learning
J. Hare · 2019
Cited alongside, same era.
A review: On path planning strategies for navigation of mobile robot
B. Patle, A. Pandey, D. Parhi, A. Jagadeesh, et al · 2019
Cited alongside, same era.
Densecavoid: Real-time navigation in dense crowds using anticipatory behaviors
A. J. Sathyamoorthy, J. Liang, U. Patel, T. Guan, R. Chandra, and D. Manocha · 2020
Cited alongside, same era.
Deep reinforcement learning in action
A. Zai and B. Brown · 2020
Cited alongside, same era.
Deep-reinforcement-learning-based autonomous uav navigation with sparse rewards
C. Wang, J. Wang, J. Wang, and X. Zhang · 2020
Cited alongside, same era.
Deep reinforcement learning for safe local planning of a ground vehicle in unknown rough terrain
S. Josef and A. Degani · 2020
Cited alongside, same era.
A review of motion planning algorithms for intelligent robots
C. Zhou, B. Huang, and P. Fränti · 2021
Later among the works it cites.
Ving: Learning open-world navigation with visual goals
D. Shah, B. Eysenbach, G. Kahn, N. Rhinehart, and S. Levine · 2021
Later among the works it cites.
T. Guan, D. Kothandaraman, R. Chandra, A. J. Sathyamoorthy, and D. Manocha · 2021
Later among the works it cites.
Cross-modal domain adaptation for cost-efficient visual reinforcement learning
X.-H. Chen, S. Jiang, F. Xu, Z. Zhang, and Y. Yu · 2021
Later among the works it cites.
On proximal policy optimization’s heavy-tailed gradients
S. Garg, J. Zhanson, E. Parisotto, A. Prasad, Z. Kolter, Z. Lipton, S. Balakrishnan, R. Salakhutdinov, and P. Ravikumar · 2021
Later among the works it cites.
Adaptiveon: Adaptive outdoor navigation method for stable and reliable motions
J. Liang, K. Weerakoon, T. Guan, N. Karapetyan, and D. Manocha · 2022
Closest in time.
Dealing with sparse rewards in continuous control robotics via heavy-tailed policies
S. Chakraborty, A. S. Bedi, A. Koppel, P. Tokekar, and D. Manocha · 2022
Closest in time.
Terrapn: Unstructured terrain navigation through online self-supervised learning
A. J. Sathyamoorthy, K. Weerakoon, T. Guan, J. Liang, and D. Manocha · 2022
Closest in time.
Terp: Reliable planning in uneven outdoor environments using deep reinforcement learning
K. Weerakoon, A. J. Sathyamoorthy, U. Patel, and D. Manocha · 2022
Closest in time.
Motion planning and control for mobile robot navigation using machine learning: a survey
X. Xiao, B. Liu, G. Warnell, and P. Stone · 2022
Closest in time.
On the hidden biases of policy mirror ascent in continuous action spaces
A. S. Bedi, S. Chakraborty, A. Parayil, B. M. Sadler, P. Tokekar, and A. Koppel · 2022
Closest in time.