Fetching the paper…
Reading the bibliography…
We introduce the first end-to-end Deep Reinforcement Learning (DRL) based framework for active high frequency trading in the stock market.
J. Močkus, “On bayesian methods for seeking the extremum,” in Optimization techniques IFIP technical conference . Springer, 1975, pp. 400–404
1975
Earlier work this paper cites.
R. Neuneier, “Optimal asset allocation using adaptive dynamic programming,” in Advances in Neural Information Processing Systems , 1996, pp. 952–958
1996
Earlier work this paper cites.
J. Moody, L. Wu, Y. Liao, and M. Saffell, “Performance functions and reinforcement learning for trading systems and portfolios,” Journal of Forecasting , vol. 17, no. 5-6, pp. 441–470, 1998
1998
Earlier work this paper cites.
D. R. Jones, M. Schonlau, and W. J. Welch, “Efficient global optimization of expensive black-box functions,” Journal of Global optimization , vol. 13, no. 4, pp. 455–492, 1998
1998
Earlier work this paper cites.
J. E. Moody and M. Saffell, “Reinforcement learning for trading,” in Advances in Neural Information Processing Systems , 1999, pp. 917–923
1999
Earlier work this paper cites.
J. Moody, M. Saffell, W. L. Andrew, Y. S. Abu-Mostafa, B. LeBaraon, and A. S. Weigend, “Minimizing downside risk via stochastic dynamic programming,” Computational Finance , pp. 403–415, 1999
1999
Earlier work this paper cites.
J. Moody and M. Saffell, “Learning to trade via direct reinforcement,” IEEE transactions on neural Networks , vol. 12, no. 4, pp. 875–889, 2001
2001
Earlier work this paper cites.
O. Mihatsch and R. Neuneier, “Risk-sensitive reinforcement learning,” Machine learning , vol. 49, no. 2-3, pp. 267–290, 2002
2002
Earlier work this paper cites.
C. E. Rasmussen, “Gaussian processes in machine learning,” in Summer School on Machine Learning . Springer, 2003, pp. 63–71
2003
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in Proceedings of the 28th International Conference on machine learning (ICML-11) , 2011, pp. 465–472
2011
Earlier work this paper cites.
F. Hutter, H. H. Hoos, and K. Leyton-Brown, “Sequential model-based optimization for general algorithm configuration,” in International conference on learning and intelligent optimization . Springer, 2011, pp. 507–523
2011
Earlier work this paper cites.
R. Huang and T. Polak, “Lobster: Limit order book reconstruction system,” Available at SSRN 1977207 , 2011
2011
Earlier work this paper cites.
F. Bertoluzzo and M. Corazza, “Testing different reinforcement learning configurations for financial trading: Introduction and applications,” Procedia Economics and Finance , vol. 3, pp. 68–77, 2012
2012
Earlier work this paper cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,” The International Journal of Robotics Research , vol. 32, no. 11, pp. 1238–1274, 2013
2013
Earlier work this paper cites.
M. Kearns and Y. Nevmyvaka, “Machine learning for market microstructure and high frequency trading,” High Frequency Trading: New Realities for Traders, Markets, and Regulators , 2013
2013
Earlier work this paper cites.
C. Comerton-Forde and T. J. Putniņš, “Dark trading and price discovery,” Journal of Financial Economics , vol. 118, no. 1, pp. 70–92, 2015
2015
Earlier work this paper cites.
M. O’Hara, “High frequency market microstructure,” Journal of Financial Economics , vol. 116, no. 2, pp. 257–270, 2015
2015
Earlier work this paper cites.
B. Shahriari, K. Swersky, Z. Wang, R. P. Adams, and N. De Freitas, “Taking the human out of the loop: A review of bayesian optimization,” Proceedings of the IEEE , vol. 104, no. 1, pp. 148–175, 2015
2015
Earlier work this paper cites.
P. Abbeel and J. Schulman, “Deep reinforcement learning through policy optimization,” Tutorial at Neural Information Processing Systems , 2016
2016
Cited alongside, same era.
Y. Deng, F. Bao, Y. Kong, Z. Ren, and Q. Dai, “Deep direct reinforcement learning for financial signal representation and trading,” IEEE transactions on neural networks and learning systems , vol. 28, no. 3, pp. 653–664, 2016
2016
Cited alongside, same era.
M. Bibinger, M. Jirak, M. Reiss et al. , “Volatility estimation under one-sided errors with applications to limit order books,” The Annals of Applied Probability , vol. 26, no. 5, pp. 2754–2790, 2016
2016
Cited alongside, same era.
Y. Ling, S. A. Hasan, V. Datla, A. Qadir, K. Lee, J. Liu, and O. Farri, “Diagnostic inferencing via improving clinical concept extraction with deep reinforcement learning: A preliminary study,” in Machine Learning for Healthcare Conference , 2017, pp. 271–285
2017
Cited alongside, same era.
T. L. Meng and M. Khushi, “Reinforcement learning in financial markets,” Data , vol. 4, no. 3, p. 110, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
L. Chen and Q. Gao, “Application of deep reinforcement learning on automated stock trading,” in 2019 IEEE 10th International Conference on Software Engineering and Service Science (ICSESS) . IEEE, 2019, pp. 29–33
2019
Later among the works it cites.
Q.-V. Dang, “Reinforcement learning in stock trading,” in International Conference on Computer Science, Applied Mathematics and Applications . Springer, 2019, pp. 311–322
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Li, “Deep reinforcement learning: An overview,” arXiv preprint arXiv:1701.07274 , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Cited alongside, same era.
A. Tsantekidis, N. Passalis, A. Tefas, J. Kanniainen, M. Gabbouj, and A. Iosifidis, “Forecasting stock prices from the limit order book using convolutional neural networks,” in 2017 IEEE 19th Conference on Business Informatics (CBI) , vol. 1. IEEE, 2017, pp. 7–12
2017
Cited alongside, same era.
Y. Kim, W. Ahn, K. J. Oh, and D. Enke, “An intelligent hybrid trading system for discovering trading rules for the futures market using rough sets and genetic algorithms,” Applied Soft Computing , vol. 55, pp. 127–140, 2017
2017
Cited alongside, same era.
Z. Jiang and J. Liang, “Cryptocurrency portfolio management with deep reinforcement learning,” in 2017 Intelligent Systems Conference (IntelliSys) . IEEE, 2017, pp. 905–913
2017
Cited alongside, same era.
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” https://github.com/hill-a/stable-baselines
2018
Cited alongside, same era.
J.-P. Bouchaud, J. Bonart, J. Donier, and M. Gould, Trades, Quotes and Prices: Financial Markets Under the Microscope . Cambridge University Press, 2018
2018
Cited alongside, same era.
G. Jeong and H. Y. Kim, “Improving financial trading decisions using deep q-learning: Predicting the number of shares, action strategies, and transfer learning,” Expert Systems with Applications , vol. 117, pp. 125–138, 2019
2019
Later among the works it cites.
Y.-J. Hu and S.-J. Lin, “Deep reinforcement learning for optimizing finance portfolio management,” in 2019 Amity International Conference on Artificial Intelligence (AICAI) . IEEE, 2019, pp. 14–20
2019
Later among the works it cites.
Y. Wang, “Electronic market making on large tick assets,” Ph.D. dissertation, The Chinese University of Hong Kong (Hong Kong), 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Bibinger, C. Neely, and L. Winkelmann, “Estimation of the discontinuous leverage effect: Evidence from the nasdaq order book,” Journal of Econometrics , vol. 209, no. 2, pp. 158–184, 2019
2019
Later among the works it cites.
F. Archetti and A. Candelieri, Bayesian Optimization and Data Science . Springer, 2019
2019
Later among the works it cites.
A. Candelieri and F. Archetti, “Global optimization in machine learning: the design of a predictive analytics application,” Soft Computing , vol. 23, no. 9, pp. 2969–2977, 2019
2019
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Z. Zhang, S. Zohren, and S. Roberts, “Deep reinforcement learning for trading,” The Journal of Financial Data Science , vol. 2, no. 2, pp. 25–40, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
H. Yang, X.-Y. Liu, S. Zhong, and A. Walid, “Deep reinforcement learning for automated stock trading: An ensemble strategy,” Available at SSRN , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.