Fetching the paper…
Reading the bibliography…
Although it is well known that exploration plays a key role in Reinforcement Learning (RL), prevailing exploration strategies for continuous control tasks in RL are mainly based on naive isotropic Gaussian noise regardless of the causality relationship between action space and the task and consider all dimensions of actions equally important.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Actor-critic algorithms
Vijay R Konda and John N Tsitsiklis · 2000
Earlier work this paper cites.
Action elimination and stopping conditions for the multi-armed bandit and reinforcement learning problems
Eyal Even-Dar, Shie M., et al · 2006
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, et al · 2009
Earlier work this paper cites.
Deterministic policy gradient algorithms
David Silver, Guy Lever, et al · 2014
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, et al · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, et al · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Martín Abadi, Paul Barham, Jianmin Chen, et al · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, et al · 2016
Earlier work this paper cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, et al · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, et al · 2017
Cited alongside, same era.
Learning to explain: An information-theoretic perspective on model interpretation
Jianbo Chen, Le Song, et al · 2018
Cited alongside, same era.
Mix&match-agent curricula for reinforcement learning
Wojciech Marian Czarnecki, Siddhant M Jayakumar, Max Jaderberg, et al · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke Van Hoof, and David Meger · 2018
Cited alongside, same era.
David Ha and Jürgen Schmidhuber · 2018
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
Danijar Hafner, Timothy Lillicrap, Jimmy Ba, et al · 2019
Later among the works it cites.
When to trust your model: Model-based policy optimization
Michael Janner, Justin Fu, Marvin Zhang, et al · 2019
Later among the works it cites.
Benchmarking model-based reinforcement learning
Eric Langlois, Shunshi Zhang, Guodong Zhang, et al · 2019
Later among the works it cites.
Teacher-student curriculum learning
Tambet Matiisen, Avital Oliver, et al · 2019
Later among the works it cites.
Towards interpretable reinforcement learning using attention augmented agents
Alexander Mott, Daniel Zoran, et al · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, et al · 2018
Cited alongside, same era.
Tstarbots: Defeating the cheating level builtin ai in starcraft ii in the full game
Peng Sun, Xinghai Sun, Lei Han, et al · 2018
Cited alongside, same era.
Curriculum learning by transfer learning: Theory and experiments with deep networks
Daphna Weinshall, Gad Cohen, et al · 2018
Cited alongside, same era.
Invase: Instance-wise variable selection using neural networks
Jinsung Yoon, James Jordon, and Mihaela van der Schaar · 2018
Cited alongside, same era.
Learn what not to learn: Action elimination with deep reinforcement learning
Tom Zahavy, Matan Haroush, et al · 2018
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, et al · 2019
Cited alongside, same era.
Causal confusion in imitation learning
Pim de Haan, Dinesh Jayaraman, and Sergey Levine · 2019
Cited alongside, same era.
Archit Sharma, Shixiang Gu, Sergey Levine, et al · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, et al · 2019
Later among the works it cites.
Learning dexterous in-hand manipulation
OpenAI: Marcin Andrychowicz, Bowen Baker, Maciek Chociej, et al · 2020
Later among the works it cites.
Instance-wise feature grouping
Aria Masoomi, Chieh Wu, Tingting Zhao, et al · 2020
Later among the works it cites.
Neuroevolution of self-interpretable agents
Yujin Tang, Duong Nguyen, and David Ha · 2020
Later among the works it cites.
What went wrong and when? instance-wise feature importance for time-series black-box models
Sana Tonekaboni, S. Joshi, et al · 2020
Later among the works it cites.
Curriculum learning for natural language understanding
Benfeng Xu, L. Zhang, et al · 2020
Later among the works it cites.