Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) has shown great success in solving many challenging tasks via use of deep neural networks.
Efficient training of artificial neural networks for autonomous navigation
Dean A Pomerleau. 1991 · 1991
Earlier work this paper cites.
On integrating apprentice learning and reinforcement learning
Jeffery Allen Clouse. 1996 · 1996
Earlier work this paper cites.
Policy invariance under reward transformations: Theory and application to reward shaping. In ICML
A Ng, D Harada, and S Russell. 1999 · 1999
Earlier work this paper cites.
Is imitation learning the route to humanoid robots?
Stefan Schaal. 1999 · 1999
Earlier work this paper cites.
Coordination in Multiagent Reinforcement Learning: A Bayesian Approach. In Proceedings of the Second International Joint Conference on Autonomous Agents and Multiagent Systems (Melbourne, Australia) (AAMAS ’03) . Association for Computing Machinery, New York, NY, USA, 709–716
Georgios Chalkiadakis and Craig Boutilier. 2003 · 2003
Earlier work this paper cites.
A survey of robot learning from demonstration
Brenna D Argall, Sonia Chernova, Manuela Veloso, and Brett Browning. 2009 · 2009
Earlier work this paper cites.
Student-Initiated Action Advising via Advice Novelty
Ercument Ilhan and Diego Perez Liebana. 2020 · 2010
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2013 · 2013
Earlier work this paper cites.
Policy shaping: Integrating human feedback with reinforcement learning. In NIPS
Shane Griffith, Kaushik Subramanian, Jonathan Scholz, Charles Isbell, and Andrea L Thomaz. 2013 · 2013
Earlier work this paper cites.
Teaching on a budget: Agents advising agents in reinforcement learning. In Proceedings of the 2013 international conference on Autonomous agents and multi-agent systems . 1053–1060
Lisa Torrey and Matthew Taylor. 2013 · 2013
Cited alongside, same era.
Reinforcement learning agents providing advice in complex video games
Matthew E Taylor, Nicholas Carboni, Anestis Fachantidis, Ioannis Vlahavas, and Lisa Torrey. 2014 · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Interactive Teaching Strategies for Agent Training. In In Proceedings of IJCAI 2016
Ofra Amir, Ece Kamar, Andrey Kolobov, and Barbara Grosz. 2016 · 2016
Cited alongside, same era.
Active Advice Seeking for Inverse Reinforcement Learning. In Proceedings of the 2016 International Conference on Autonomous Agents & Multiagent Systems, Singapore, May 9-13, 2016 . 512–520
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 2018 · 2018
Later among the works it cites.
Learning the dynamic treatment regimes from medical registry data through deep Q-network
Ning Liu, Ying Liu, Brent Logan, Zhiyuan Xu, Jian Tang, and Yanzhi Wang. 2019 · 2019
Later among the works it cites.
Learning to teach in cooperative multiagent reinforcement learning. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 6128–6136
Shayegan Omidshafiei, Dong-Ki Kim, Miao Liu, Gerald Tesauro, Matthew Riemer, Christopher Amato, Murray Campbell, and Jonathan P How. 2019 · 2019
Later among the works it cites.
Optimization of molecules via deep reinforcement learning
Zhenpeng Zhou, Steven Kearnes, Li Li, Richard N Zare, and Patrick Riley. 2019 · 2019
Later among the works it cites.
Uncertainty-aware action advising for deep reinforcement learning agents. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 34. 5792–5799
Felipe Leno Da Silva, Pablo Hernandez-Leal, Bilal Kartal, and Matthew E Taylor. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Phillip Odom and Sriraam Natarajan. 2016 · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Agent-aware dropout dqn for safe and efficient on-line dialogue policy learning. In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing . 2454–2464
Lu Chen, Xiang Zhou, Cheng Chang, Runzhe Yang, and Kai Yu. 2017 · 2017
Cited alongside, same era.
Simultaneously learning and advising in multiagent reinforcement learning. In Proceedings of the 16th conference on autonomous agents and multiagent systems . 1100–1108
Felipe Leno Da Silva, Ruben Glatt, and Anna Helena Reali Costa. 2017 · 2017
Cited alongside, same era.
Generalization and Regularization in DQN
Jesse Farebrother, Marlos C. Machado, and Michael Bowling. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Learning by reusing previous advice in teacher-student paradigm. In Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems . 1674–1682
Changxi Zhu, Yi Cai, Ho-fung Leung, and Shuyue Hu. 2020 · 2020
Later among the works it cites.
Towered Actor Critic For Handling Multiple Action Types In Reinforcement Learning For Drug Discovery. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 142–150
Sai Krishna Gottipati, Yashaswi Pathak, Boris Sattarov, Sahir, Rohan Nuttall, Mohammad Amini, Matthew E.T̃aylor, and Sarath Chandar. 2021 · 2021
Later among the works it cites.
Action Advising with Advice Imitation in Deep Reinforcement Learning. In AAMAS ’21: 20th International Conference on Autonomous Agents and Multiagent Systems, Virtual Event, United Kingdom, May 3-7, 2021 , Frank Dignum, Alessio Lomuscio, Ulle Endriss, and Ann Nowé (Eds.). ACM, 629–637
Ercument Ilhan, Jeremy Gow, and Diego Perez Liebana. 2021a · 2021
Later among the works it cites.
Learning on a Budget via Teacher Imitation
Ercument Ilhan, Jeremy Gow, and Diego Perez Liebana. 2021b · 2021
Later among the works it cites.