Fetching the paper…
Reading the bibliography…
Supervised learning, more specifically Convolutional Neural Networks (CNN), has surpassed human ability in some visual recognition tasks such as detection of traffic signs, faces and handwritten numbers.
L. J. Lin, “Self-improving reactive agents based on reinforcement learning, planning and teaching,” Machine Learning , vol. 8, pp. 293–321, 1992
1992
Earlier work this paper cites.
N. Kohl and P. Stone, “Policy gradient reinforcement learning for fast quadrupedal locomotion,” in IEEE International Conference on Robotics and Automation, 2004. Proceedings. ICRA ’04. 2004 , vol. 3, April 2004, pp. 2619–2624 Vol.3
2004
Earlier work this paper cites.
S. Cheng and L. M. Frank, “New experiences enhance coordinated neural activity in the hippocampus,” Neuron , vol. 57, no. 2, pp. 303 – 313, 2008. [Online]. Available: http://www.sciencedirect.com/science/article/pii/S0896627307010306
2008
Earlier work this paper cites.
G. Girardeau, K. Benchenane, S. I. Wiener, G. Buzsáki, and M. B. Zugaro, “Selective suppression of hippocampal ripples impairs spatial memory,” Nature Neuroscience , vol. 12, pp. 1222 EP –, Sep 2009. [Online]. Available: http://dx.doi.org/10.1038/nn.2384
2009
Earlier work this paper cites.
E. Theodorou, J. Buchli, and S. Schaal, “Reinforcement learning of motor skills in high dimensions: A path integral approach,” in 2010 IEEE International Conference on Robotics and Automation , May 2010, pp. 2397–2403
2010
Earlier work this paper cites.
V. Ego-Stengel and M. A. Wilson, “Disruption of ripple-associated hippocampal activity during rest impairs spatial learning in the rat,” Hippocampus , vol. 20, no. 1, pp. 1–10, Jan 2010, 19816984[pmid]. [Online]. Available: http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2801761/
2010
Earlier work this paper cites.
M. Kalakrishnan, L. Righetti, P. Pastor, and S. Schaal, “Learning force control policies for compliant manipulation,” in 2011 IEEE/RSJ International Conference on Intelligent Robots and Systems , Sept 2011, pp. 4639–4644
2011
Earlier work this paper cites.
M. F. Carr, S. P. Jadhav, and L. M. Frank, “Hippocampal replay in the awake state: a potential substrate for memory consolidation and retrieval,” Nature Neuroscience , vol. 14, pp. 147 EP –, Jan 2011, review Article. [Online]. Available: http://dx.doi.org/10.1038/nn.2732
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in Intelligent Robots and Systems (IROS), 2012 IEEE/RSJ International Conference on . IEEE, 2012, pp. 5026–5033
2012
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing atari with deep reinforcement learning,” in NIPS Deep Learning Workshop , 2013
2013
Cited alongside, same era.
2015
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, Feb. 2015. [Online]. Available: http://dx.doi.org/10.1038/nature14236
2016
Later among the works it cites.
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, O. P. Abbeel, and W. Zaremba, “Hindsight experience replay,” in Advances in Neural Information Processing Systems , 2017, pp. 5048–5058
2017
Later among the works it cites.
T. Le, S. Gibb, N. Pham, H. M. La, L. Falk, and T. Berendsen, “Autonomous robotic system using non-destructive evaluation methods for bridge deck inspection,” in Robotics and Automation (ICRA), 2017 IEEE International Conference on . IEEE, 2017, pp. 3672–3677
2017
Later among the works it cites.
S. Gibb, T. Le, H. M. La, R. Schmid, and T. Berendsen, “A multi-functional inspection robot for civil infrastructure evaluation and maintenance,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , Sept 2017, pp. 2672–2677
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
L. A. Atherton, D. Dupret, and J. R. Mellor, “Memory trace replay: the shaping of memory consolidation by neuromodulation,” Trends in Neurosciences , vol. 38, no. 9, pp. 560 – 570, 2015. [Online]. Available: http://www.sciencedirect.com/science/article/pii/S0166223615001551
2015
Cited alongside, same era.
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine, “Continuous deep q-learning with model-based acceleration,” in Proceedings of The 33rd International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, M. F. Balcan and K. Q. Weinberger, Eds., vol. 48. New York, New York, USA: PMLR, 20–22 Jun 2016, pp. 2829–2838. [Online]. Available: http://proceedings.mlr.press/v48/gu16.html
2016
Cited alongside, same era.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016
2016
Cited alongside, same era.
T. Schaul, J. Quan, I. Antonoglou, and D. Silver, “Prioritized experience replay,” in International Conference on Learning Representations , Puerto Rico, 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Closest in time.
S. Gibb, H. M. La, T. Le, L. Nguyen, R. Schmid, and H. Pham, “Nondestructive evaluation sensor fusion with autonomous robotic system for civil infrastructure inspection,” Journal of Field Robotics , vol. 0, no. 0, 2018. [Online]. Available: https://onlinelibrary.wiley.com/doi/abs/10.1002/rob.21791
2018
Closest in time.