Fetching the paper…
Reading the bibliography…
As a fundamental problem in algorithmic trading, order execution aims at fulfilling a specific trading order, either liquidation or acquirement, for a given instrument.
Distillation Strategies for Proximal Policy Optimization
Green, S.; Vineyard, C. M.; and Koç, C. K. 2019 · 1901
Earlier work this paper cites.
Optimization of trading systems and portfolios
Moody, J.; and Wu, L. 1997 · 1997
Earlier work this paper cites.
Optimal control of execution costs
Bertsimas, D.; and Lo, A. W. 1998 · 1998
Earlier work this paper cites.
Reinforcement Learning for Trading Systems and Portfolios
Moody, J. E.; Saffell, M.; Liao, Y.; and Wu, L. 1998 · 1998
Earlier work this paper cites.
Optimal execution of portfolio transactions
Almgren, R.; and Chriss, N. 2001 · 2001
Earlier work this paper cites.
Learning to trade via direct reinforcement
Moody, J.; and Saffell, M. 2001 · 2001
Earlier work this paper cites.
Areas beneath the relative operating characteristics (ROC) and relative operating levels (ROL) curves: Statistical significance and interpretation
Mason, S. J.; and Graham, N. E. 2002 · 2002
Earlier work this paper cites.
Conditional Mutual information-based Contrastive Loss for Financial Time Series Forecasting
Wu, H.; Gattami, A.; and Flierl, M. 2020 · 2002
Earlier work this paper cites.
Reinforcement-learning based portfolio management with augmented asset movement prediction states
Ye, Y.; Pei, H.; Wang, B.; Chen, P.-Y.; Zhu, Y.; Xiao, J.; Li, B.; et al. 2020 · 2002
Earlier work this paper cites.
Suphx: Mastering Mahjong with Deep Reinforcement Learning
Li, J.; Koyamada, S.; Ye, Q.; Liu, G.; Wang, C.; Yang, R.; Zhao, L.; Qin, T.; Liu, T.-Y.; and Hon, H.-W. 2020 · 2003
Earlier work this paper cites.
Competitive algorithms for VWAP and limit order trading
Kakade, S. M.; Kearns, M.; Mansour, Y.; and Ortiz, L. E. 2004 · 2004
Earlier work this paper cites.
Lai, K.-H.; Zha, D.; Li, Y.; and Hu, X. 2020 · 2006
Earlier work this paper cites.
Reinforcement learning for optimized trade execution
Nevmyvaka, Y.; Feng, Y.; and Kearns, M. 2006 · 2006
Earlier work this paper cites.
Improving VWAP strategies: A dynamic volume approach
Białkowski, J.; Darolles, S.; and Le Fol, G. 2008 · 2008
Cited alongside, same era.
A reinforcement learning extension to the Almgren-Chriss framework for optimal trade execution
Hendricks, D.; and Wilcox, D. 2014 · 2014
Cited alongside, same era.
Optimal execution with limit and market orders
Cartea, A.; and Jaimungal, S. 2015 · 2015
Cited alongside, same era.
Algorithmic and high-frequency trading
Cartea, Á.; Jaimungal, S.; and Penalva, J. 2015 · 2015
Cited alongside, same era.
General intensity shapes in optimal liquidation
Guéant, O.; and Lehalle, C.-A. 2015 · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Hinton, G.; Vinyals, O.; and Dean, J. 2015 · 2015
Cryptocurrency portfolio management with deep reinforcement learning
Jiang, Z.; and Liang, J. 2017 · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Schulman, J.; Wolski, F.; Dhariwal, P.; Radford, A.; and Klimov, O. 2017 · 2017
Later among the works it cites.
Distral: Robust multitask reinforcement learning
Teh, Y.; Bapst, V.; Czarnecki, W. M.; Quan, J.; Kirkpatrick, J.; Hadsell, R.; Heess, N.; and Pascanu, R. 2017 · 2017
Later among the works it cites.
Knowledge transfer for deep reinforcement learning with hierarchical experience replay
Yin, H.; and Pan, S. J. 2017 · 2017
Later among the works it cites.
Double Deep Q-Learning for Optimal Execution
Ning, B.; Ling, F. H. T.; and Jaimungal, S. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Cited alongside, same era.
Actor-mimic: Deep multitask and transfer reinforcement learning
Parisotto, E.; Ba, J. L.; and Salakhutdinov, R. 2015 · 2015
Cited alongside, same era.
Rusu, A. A.; Colmenarejo, S. G.; Gulcehre, C.; Desjardins, G.; Kirkpatrick, J.; Pascanu, R.; Mnih, V.; Kavukcuoglu, K.; and Hadsell, R. 2015 · 2015
Cited alongside, same era.
Incorporating order-flow into optimal execution
Cartea, A.; and Jaimungal, S. 2016 · 2016
Cited alongside, same era.
The Financial Mathematics of Market Liquidity: From optimal execution to market making , volume 33
Guéant, O. 2016 · 2016
Cited alongside, same era.
Optimal Order Execution using Stochastic Control and Reinforcement Learning
Hu, R. 2016 · 2016
Cited alongside, same era.
Reinforcement learning: An introduction
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Later among the works it cites.
Trading algorithms with learning in latent alpha models
Casgrain, P.; and Jaimungal, S. 2019 · 2019
Later among the works it cites.
Deep Execution-Value and Policy Based Reinforcement Learning for Trading and Beating Market Benchmarks
Dabérius, K.; Granat, E.; and Karlsson, P. 2019 · 2019
Later among the works it cites.
A cooperative multi-agent reinforcement learning framework for resource balancing in complex logistics network
Li, X.; Zhang, J.; Bian, J.; Tong, Y.; and Liu, T.-Y. 2019 · 2019
Later among the works it cites.
A deep value-network based approach for multi-driver order dispatching
Tang, X.; Qin, Z.; Zhang, F.; Wang, Z.; Xu, Z.; Ma, Y.; Zhu, H.; and Ye, J. 2019 · 2019
Later among the works it cites.
CSI800 index
Co., C. S. I. 2020 · 2020
Later among the works it cites.
Reinforcement Learning with Non-Markovian Rewards
Gaon, M.; and Brafman, R. 2020 · 2020
Later among the works it cites.
An End-to-End Optimal Trade Execution Framework based on Proximal Policy Optimization
Lin, S.; and Beling, P. A. 2020 · 2020
Later among the works it cites.