Fetching the paper…
Reading the bibliography…
In this paper, we implement three state-of-art continuous reinforcement learning algorithms, Deep Deterministic Policy Gradient (DDPG), Proximal Policy Optimization (PPO) and Policy Gradient (PG)in portfolio management.
Gao, Xiu, and Laiwan Chan. ”An algorithm for trading and portfolio management using Q-learning and sharpe ratio maximization.” Proceedings of the international conference on neural information processing. 2000
2000
Earlier work this paper cites.
Sutton, Richard S., et al. ”Policy gradient methods for reinforcement learning with function approximation.” Advances in neural information processing systems. 2000
2000
Earlier work this paper cites.
Jacobsen, Ben, and Dennis Dannenburg. ”Volatility clustering in monthly stock returns.” Journal of Empirical Finance 10.4 (2003): 479-503
2003
Earlier work this paper cites.
Du, Xin, Jinjian Zhai, and Koupin Lv. ”Algorithm trading using q-learning and recurrent reinforcement learning.” positions 1 (2009): 1
2009
Earlier work this paper cites.
Rua, António, and Luís C. Nunes. ”International comovement of stock market returns: A wavelet analysis.” Journal of Empirical Finance 16.4 (2009): 632-639
2009
Earlier work this paper cites.
Rao, Anil V. ”A survey of numerical methods for optimal control.” Advances in the Astronautical Sciences 135.1 (2009): 497-528
2009
Earlier work this paper cites.
Faragher, Ramsey. ”Understanding the basis of the Kalman filter via a simple and intuitive derivation.” IEEE Signal processing magazine 29.5 (2012): 128-132
2012
Earlier work this paper cites.
Prashanth, L. A., and Mohammad Ghavamzadeh. ”Actor-critic algorithms for risk-sensitive MDPs.” Advances in neural information processing systems. 2013
2013
Earlier work this paper cites.
Li, Bin, and Steven CH Hoi. ”Online portfolio selection: A survey.” ACM Computing Surveys (CSUR) 46.3 (2014): 35
2014
Earlier work this paper cites.
Silver, David, et al. ”Deterministic policy gradient algorithms.” ICML. 2014
2014
Cited alongside, same era.
Dewey, Daniel. ”Reinforcement learning and the reward engineering principle.” 2014 AAAI Spring Symposium Series. 2014
2014
Cited alongside, same era.
Mnih, Volodymyr, et al. ”Human-level control through deep reinforcement learning.” Nature 518.7540 (2015): 529
2015
Cited alongside, same era.
2015
Cited alongside, same era.
He K, Zhang X, Ren S, et al. Deep residual learning for image recognition[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2016: 770-778
2016
Cited alongside, same era.
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gu, Shixiang, et al. ”Continuous deep q-learning with model-based acceleration.” International Conference on Machine Learning. 2016
2016
Cited alongside, same era.
Almahdi, Saud, and Steve Y. Yang. ”An adaptive portfolio trading system: A risk-return portfolio optimization using recurrent reinforcement learning with expected maximum drawdown.” Expert Systems with Applications 87 (2017): 267-279
2017
Cited alongside, same era.
Li, Yuxi. ”Deep reinforcement learning: An overview.” arXiv preprint arXiv:1701.07274 (2017)
2017
Cited alongside, same era.
2017
Cited alongside, same era.
John Schulman, Sergey Levine, Philipp Moritz, Michael Jordan, Pieter Abbel: Trust Region Policy Optimization
Cited in the paper.
Tang, Lili. ”An actor-critic-based portfolio investment method inspired by benefit-risk optimization.” Journal of Algorithms and Computational Technology (2018): 1748301818779059
2018
Closest in time.
Yang, Steve Y., Yangyang Yu, and Saud Almahdi. ”An investor sentiment reward-based trading system using Gaussian inverse reinforcement learning algorithm.” Expert Systems with Applications 114 (2018): 388-401
2018
Closest in time.
Buehler, Hans, et al. ”Deep hedging.” (2018)
2018
Closest in time.
Pattanaik, Anay, et al. ”Robust deep reinforcement learning with adversarial attacks.” Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems. International Foundation for Autonomous Agents and Multiagent Systems, 2018
2018
Closest in time.