CarRacing-v0
O. Klimov · 2016
Later among the works it cites.
Using Keras and deep deterministic policy gradient to play TORCS
B. Lau · 2016
Later among the works it cites.
DoomTakeCover-v0
P. Paquette · 2016
Later among the works it cites.
The predictron: End-to-end learning and planning
Original
D. Silver, H. van Hasselt, M. Hessel, T. Schaul, A. Guez, T. Harley, G. Dulac-Arnold, D. Reichert, N. Rabinowitz, A. Barreto, and T. Degris · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Original
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, and K. Kavukcuoglu · 2016
Later among the works it cites.
Autoencoder-augmented neuroevolution for visual doom playing
S. Alvernaz and J. Togelius · 2017
Later among the works it cites.
Using simulation and domain adaptation to improve efficiency of deep robotic grasping
Original
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige, S. Levine, and V. Vanhoucke · 2017
Later among the works it cites.
The code for facial identity in the primate brain
L. Chang and D. Y. Tsao · 2017
Later among the works it cites.
Recurrent environment simulators
Original
S. Chiappa, S. Racaniere, D. Wierstra, and S. Mohamed · 2017
Later among the works it cites.
Unsupervised learning of disentangled representations from video
E. L. Denton et al · 2017
Later among the works it cites.
Pathnet: Evolution channels gradient descent in super neural networks
Original
C. Fernando, D. Banarse, C. Blundell, Y. Zwols, D. Ha, A. Rusu, A. Pritzel, and D. Wierstra · 2017
Later among the works it cites.
Generative temporal models with memory
Original
M. Gemici, C. Hung, A. Santoro, G. Wayne, S. Mohamed, D. Rezende, D. Amos, and T. Lillicrap · 2017
Later among the works it cites.
Evolving stable strategies
D. Ha · 2017
Later among the works it cites.
Hypernetworks
D. Ha, A. Dai, and Q. V. Le · 2017
Later among the works it cites.
A benchmark environment motivated by industrial control problems
Original
D. Hein, S. Depeweg, M. Tokic, S. Udluft, A. Hentschel, T. Runkler, and V. Sterzing · 2017
Later among the works it cites.
DARLA: Improving zero-shot transfer in reinforcement learning
Original
I. Higgins, A. Pal, A. A. Rusu, L. Matthey, C. P. Burgess, A. Pritzel, M. Botvinick, C. Blundell, and A. Lerchner · 2017
Later among the works it cites.
Self-driving cars in the browser
J. Hünermann · 2017
Later among the works it cites.
Reinforcement car racing with A3C
S. Jang, J. Min, and C. Lee · 2017
Later among the works it cites.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska, et al · 2017
Later among the works it cites.
A sensorimotor circuit in mouse cortex for visual flow predictions
M. Leinweber, D. R. Ward, J. M. Sobczak, A. Attinger, and G. B. Keller · 2017
Later among the works it cites.
Game engine learning from video
M. O. R. Matthew Guzdial, Boyang Li · 2017
Later among the works it cites.
Data-efficient reinforcement learning in continuous state-action Gaussian-POMDPs
R. McAllister and C. E. Rasmussen · 2017
Later among the works it cites.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Original
A. Nagabandi, G. Kahn, R. Fearing, and S. Levine · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Later among the works it cites.
Deep-Q Learning for racecar reinforcement learning problem
L. Prieur · 2017
Later among the works it cites.
Imagination-augmented agents for deep reinforcement learning
S. Racanière, T. Weber, D. Reichert, L. Buesing, A. Guez, D. J. Rezende, A. P. Badia, O. Vinyals, N. Heess, Y. Li, et al · 2017
Later among the works it cites.
Evolution strategies as a scalable alternative to reinforcement learning
Original
T. Salimans, J. Ho, X. Chen, S. Sidor, and I. Sutskever · 2017
Later among the works it cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
N. Shazeer, A. Mirhoseini, K. Maziarz, A. Davis, Q. Le, G. Hinton, and J. Dean · 2017
Later among the works it cites.
Language modeling with recurrent highway hypernetworks
J. Suarez · 2017
Later among the works it cites.
Deep neuroevolution: Genetic algorithms are a competitive alternative for training deep neural networks for reinforcement learning
Original
F. P. Such, V. Madhavan, E. Conti, J. Lehman, K. O. Stanley, and J. Clune · 2017
Later among the works it cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Visual interaction networks
Original
N. Watters, A. Tacchetti, T. Weber, R. Pascanu, P. Battaglia, and D. Zoran · 2017
Later among the works it cites.
A neural representation of sketch drawings
D. Ha and D. Eck · 2018
Closest in time.
One big net for everything
Original
J. Schmidhuber · 2018
Closest in time.
The Kanerva machine: A generative distributed memory
Y. Wu, G. Wayne, A. Graves, and T. Lillicrap · 2018
Closest in time.