Fetching the paper…
Reading the bibliography…
In video games, non-player characters (NPCs) are used to enhance the players' experience in a variety of ways, e.g., as enemies, allies, or innocent bystanders.
A formal basis for the heuristic determination of minimum cost paths
P. E. Hart, N. J. Nilsson, and B. Raphael · 1968
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel · 1989
Earlier work this paper cites.
Simultaneous map building and localization for an autonomous mobile robot
J. J. Leonard and H. F. Durrant-Whyte · 1991
Earlier work this paper cites.
Dyna, an integrated architecture for learning, planning, and reacting
R. S. Sutton · 1991
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Learning to Forget: Continual Prediction with LSTM
F. A. Gers, J. A. Schmidhuber, and F. A. Cummins · 2000
Earlier work this paper cites.
Simplified 3d movement and pathfinding using navigation meshes
G. Snook · 2000
Earlier work this paper cites.
The quake iii arena bot
J. Van Waveren · 2001
Earlier work this paper cites.
Strategic and tactical reasoning with waypoints
L. Lidén et al · 2002
Earlier work this paper cites.
Building a near-optimal navigation mesh
P. Tozour and I. Austin · 2002
Earlier work this paper cites.
Simultaneous localization and mapping: part i
H. Durrant-Whyte and T. Bailey · 2006
Earlier work this paper cites.
Navigating graph generation in highly dynamic worlds , volume 4 of AI Game Programming Wisdom , chapter 2.6, pages 125–141
R. Axelrod · 2008
Earlier work this paper cites.
Dynamically updating a Navigation Mesh via Efficient Polygon Subdivision , volume 4 of AI Game Programming Wisdom , chapter 2.3, pages 83–94
P. Marden · 2008
Earlier work this paper cites.
Intrinsic Detail in Navigation Mesh Generation , volume 4 of AI Game Programming Wisdom , chapter 2.4, pages 95–112
C. McAnlis · 2008
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Efficient obstacle avoidance using autonomously generated navigation meshes
S. Brand · 2009
Earlier work this paper cites.
Automatic generation of suboptimal navmeshes
R. Oliva and N. Pelechano · 2011
Cited alongside, same era.
Navigation meshes for realistic multi-layered environments
W. Van Toll, A. F. Cook, and R. Geraerts · 2011
Cited alongside, same era.
Automatic generation of jump links in arbitrary 3d environments
S. Budde · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Playing doom with slam-augmented deep reinforcement learning
S. Bhatti, A. Desmaison, O. Miksik, N. Nardelli, N. Siddharth, and P. H. Torr · 2016
Cited alongside, same era.
Unity: A general platform for intelligent agents
A. Juliani, V.-P. Berges, E. Vckay, Y. Gao, H. Henry, M. Mattar, and D. Lange · 2018
Later among the works it cites.
Recurrent experience replay in distributed reinforcement learning
S. Kapturowski, G. Ostrovski, J. Quan, R. Munos, and W. Dabney · 2018
Later among the works it cites.
Introduction to reinforcement learning
R. S. Sutton and A. G. Barto · 2018
Later among the works it cites.
Vizdoom competitions: Playing doom from pixels
M. Wydmuch, M. Kempka, and W. Jaśkowski · 2018
Later among the works it cites.
A hitchhiker’s guide to statistical comparisons of reinforcement learning algorithms
C. Colas, O. Sigaud, and P.-Y. Oudeyer · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, D. Kumaran, and R. Hadsell · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis · 2016
Cited alongside, same era.
A comparative study of navigation meshes
W. Van Toll, R. Triesscheijn, M. Kallmann, R. Oliva, N. Pelechano, J. Pettré, and R. Geraerts · 2016
Cited alongside, same era.
Hindsight experience replay
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, O. P. Abbeel, and W. Zaremba · 2017
Cited alongside, same era.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Cited alongside, same era.
Deep reinforcement learning that matters
P. Henderson, R. Islam, P. Bachman, J. Pineau, D. Precup, and D. Meger · 2017
Cited alongside, same era.
B. Eysenbach, R. R. Salakhutdinov, and S. Levine · 2019
Later among the works it cites.
Learning to reach goals without reinforcement learning
D. Ghosh, A. Gupta, J. Fu, A. Reddy, C. Devin, B. Eysenbach, and S. Levine · 2019
Later among the works it cites.
When to trust your model: Model-based policy optimization
M. Janner, J. Fu, M. Zhang, and S. Levine · 2019
Later among the works it cites.
Model-based reinforcement learning for atari
L. Kaiser, M. Babaeizadeh, P. Milos, B. Osinski, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine, R. Sepassi, G. Tucker, and H. Michalewski · 2019
Later among the works it cites.
Habitat: A Platform for Embodied AI Research
Manolis Savva, Abhishek Kadian, Oleksandr Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra · 2019
Later among the works it cites.
Scaling local control to large-scale topological navigation
X. Meng, N. Ratliff, Y. Xiang, and D. Fox · 2019
Later among the works it cites.
Making efficient use of demonstrations to solve hard exploration problems
T. L. Paine, C. Gulcehre, B. Shahriari, M. Denil, M. Hoffman, H. Soyer, R. Tanburn, S. Kapturowski, N. Rabinowitz, D. Williams, et al · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. Choi, R. Powell, T. Ewalds, P. Georgiev, J. Oh, D. Horgan, M. Kroiss, I. Danihelka, A. Huang, L. Sifre, T. Cai, J. Agapiou, M. Jaderberg, and D. Silver · 2019
Later among the works it cites.
Dd-ppo: Learning near-perfect pointgoal navigators from 2.5 billion frames
E. Wijmans, A. Kadian, A. Morcos, S. Lee, I. Essa, D. Parikh, M. Savva, and D. Batra · 2019
Later among the works it cites.
Combining optimal control and learning for visual navigation in novel environments
S. Bansal, V. Tolani, S. Gupta, J. Malik, and C. Tomlin · 2020
Closest in time.
Egomap: Projective mapping and structured egocentric memory for deep rl
E. Beeching, C. Wolf, J. Dibangoye, and O. Simonin · 2020
Closest in time.