Fetching the paper…
Reading the bibliography…
Two less addressed issues of deep reinforcement learning are (1) lack of generalization capability to new target goals, and (2) data inefficiency i.e., the model requires several (and often costly) episodes of trial and error to converge, which makes it impractical to be applied to real-world scenarios.
J. Borenstein and Y. Koren, “Real-time obstacle avoidance for fast mobile robots,”
1989
Earlier work this paper cites.
——, “The vector field histogram-fast obstacle avoidance for mobile robots,”
1991
Earlier work this paper cites.
D. Kim and R. Nevatia, “Simbolic navigation with a generic map,” in
1995
Earlier work this paper cites.
G. U. G. Oriolo and M. Vendittelli, “On-line map building and navigation for autonomous mobile robots,” in
1995
Earlier work this paper cites.
H. Haddad, M. Khatib, S. Lacroix, and R. Chatila, “Reactive navigation in outdoor environments using potential fields,” in
1998
Earlier work this paper cites.
K. Kidono, J. Miura, and Y. Shirai, “Autonomous visual navigation of a mobile robot using a human guided experience,”
2002
Earlier work this paper cites.
A. J. Davison, “Real time simultaneous localisation and mapping with a single camera,” in
2003
Earlier work this paper cites.
S. Lenser and M. Veloso, “Visual sonar: Fast obstacle avoidance using monocular vision,” in
2003
Earlier work this paper cites.
A. Remazeilles, F. Chaumette, and P. Gros, “Robot motion control from a visual memory,” in
2004
Earlier work this paper cites.
N. Kohl and P. Stone, “Policy gradient reinforcement learning for fast quadrupedal locomotion,” in
2004
Earlier work this paper cites.
H. J. Kim, M. I. Jordan, S. Sastry, and A. Y. Ng, “Autonomous helicopter flight via reinforcement learning,” in
2004
Earlier work this paper cites.
E. Royer, J. Bom, M. Dhome, B. Thuillot, M. Lhuillier, and F. Marmoiton, “Outdoor autonomous navigation using monocular vision,” in
2005
Earlier work this paper cites.
J. Michels, A. Saxena, and A. Ng, “High speed obstacle avoidance using monocular vision and reinforcement learning,” in
2005
Earlier work this paper cites.
S. Chopra, R. Hadsell, and Y. LeCun, “Learning a similarity metric discriminatively, with application to face verification,” in
2005
Earlier work this paper cites.
R. Sim and J. J. Little, “Autonomous vision-based exploration and mapping using hybrid maps and rao-blackwellised particle filters,” in
2006
Earlier work this paper cites.
D. Wooden, “A guide to vision-based map building,”
2006
Earlier work this paper cites.
M. Tomono, “3-d object map building using dense object models with sift-based recognition features,” in
2006
Earlier work this paper cites.
P. Saeedi, P. D. Lawrence, and D. G. Lowe, “Vision-based 3-d trajectory tracking for unknown environments,”
2006
Cited alongside, same era.
F. Bonin-Font, A. Ortiz, and G. Oliver, “Visual navigation for mobile robots: A survey,”
2008
Cited alongside, same era.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,”
2008
Cited alongside, same era.
T. Kollar and N. Roy, “Trajectory optimization using reinforcement learning for map exploration,”
2008
Cited alongside, same era.
L. van der Maaten and G. E. Hinton, “Visualizing data using t-sne,”
2008
Cited alongside, same era.
K. Konolige, J. Bowman, J. Chen, P. Mihelich, M. Calonder, V. Lepetit, and P. Fua, “View-based maps,”
2010
Y. Liang, M. C. Machado, E. Talvitie, and M. Bowling, “State of the art control of atari games using shallow reinforcement learning,” in
2016
Closest in time.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot,
2016
Closest in time.
A. A. Rusu, S. G. Colmenarejo, Ç. Gülçehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, and R. Hadsell, “Policy distillation,” in
2016
Closest in time.
E. Parisotto, L. J. Ba, and R. Salakhutdinov, “Actor-mimic: Deep multitask and transfer reinforcement learning,” in
2016
Closest in time.
R. Mottaghi, H. Bagherinezhad, M. Rastegari, and A. Farhadi, “Newtonian image understanding: Unfolding the dynamics of objects in static images,” in
2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,”
2013
Cited alongside, same era.
B. Wymann, E. Espié, C. Guionneau, C. Dimitrakakis, R. Coulom, and A. Sumner, “TORCS, the open racing car simulator, v1.3.5,”
2013
Cited alongside, same era.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,”
2013
Cited alongside, same era.
C. McManus, B. Upcroft, and P. Newman, “Scene signatures: Localised and point-less features for localisation,” in
2014
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski,
2015
Cited alongside, same era.
J. Wu, I. Yildirim, J. J. Lim, W. T. Freeman, and J. B. Tenenbaum, “Galileo: Perceiving physical object properties by integrating a physics engine with deep learning,” in
2015
Cited alongside, same era.
R. Mottaghi, M. Rastegari, A. Gupta, and A. Farhadi, ““what happens if…” learning to predict the effect of forces in images,” in
2016
Closest in time.
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski, “Vizdoom: A doom-based ai research platform for visual reinforcement learning,” in
2016
Closest in time.
A. Lerer, S. Gross, and R. Fergus, “Learning physical intuition of block towers by example,” in
2016
Closest in time.
M. Johnson, K. Hofmann, T. Hutton, and D. Bignell, “The malmo platform for artificial intelligence experimentation,” in
2016
Closest in time.
A. Handa, V. Patraucean, S. Stent, and R. Cipolla, “Scenenet: An annotated model generator for indoor scene understanding,” in
2016
Closest in time.
G. Ros, L. Sellart, J. Materzynska, D. Vazquez, and A. Lopez, “The SYNTHIA Dataset: A large collection of synthetic images for semantic segmentation of urban scenes,” in
2016
Closest in time.
A. Gaidon, Q. Wang, Y. Cabon, and E. Vig, “Virtual worlds as proxy for multi-object tracking analysis,” in
2016
Closest in time.
2016
Closest in time.
2016
Closest in time.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Closest in time.