Fetching the paper…
Reading the bibliography…
This paper investigates to what extent one can improve reinforcement learning algorithms.
A stochastic approximation method
Herbert Robbins and Sutton Monro · 1951
Earlier work this paper cites.
Approximation methods which converge with probability one
Julius R Blum et al · 1954
Earlier work this paper cites.
Asymptotic distribution of stochastic approximation procedures
Jerome Sacks · 1958
Earlier work this paper cites.
Functional approximations and dynamic programming
Richard Bellman and Stuart Dreyfus · 1959
Earlier work this paper cites.
Stochastic approximation methods for constrained and unconstrained systems
Harold Joseph Kushner and Dean S Clark · 1978
Earlier work this paper cites.
Algorithmes adaptatifs et approximations stochastiques: théorie et applications à l’identification, au traitement du signal et à la reconnaissance des formes
Albert Benveniste, Michel Métivier, and Pierre Priouret · 1987
Earlier work this paper cites.
Paul J Werbos · 1987
Earlier work this paper cites.
Learning from delayed rewards
Christopher John Cornish Hellaby Watkins · 1989
Earlier work this paper cites.
A nonparametric approach to pricing and hedging derivative securities via learning networks
James M Hutchinson, Andrew W Lo, and Tomaso Poggio · 1994
Earlier work this paper cites.
Neuro-dynamic programming
Dimitri P Bertsekas and John N Tsitsiklis · 1996
Earlier work this paper cites.
Introduction to reinforcement learning
Richard S Sutton, Andrew G Barto, et al · 1998
Earlier work this paper cites.
Introduction to statistical learning theory
Olivier Bousquet, Stéphane Boucheron, and Gábor Lugosi · 2003
Cited alongside, same era.
Stochastic approximation and recursive algorithms and applications
Harold Kushner and G George Yin · 2003
Cited alongside, same era.
Optimal aggregation of classifiers in statistical learning
Alexander B Tsybakov et al · 2004
Cited alongside, same era.
Convexity, classification, and risk bounds
Peter L Bartlett, Michael I Jordan, and Jon D McAuliffe · 2006
Cited alongside, same era.
Reinforcement learning for optimized trade execution
Yuriy Nevmyvaka, Yi Feng, and Michael Kearns · 2006
Cited alongside, same era.
The tradeoffs of large scale learning
Léon Bottou and Olivier Bousquet · 2008
Cited alongside, same era.
Accelerating stochastic gradient descent using predictive variance reduction
Rie Johnson and Tong Zhang · 2013
Later among the works it cites.
Saga: A fast incremental gradient method with support for non-strongly convex composite objectives
Aaron Defazio, Francis Bach, and Simon Lacoste-Julien · 2014
Later among the works it cites.
Stochastic proximal gradient descent with acceleration techniques
Atsushi Nitanda · 2014
Later among the works it cites.
Algorithmic and High-Frequency Trading (Mathematics, Finance and Risk)
Álvaro Cartea, Sebastian Jaimungal, and José Penalva · 2015
Later among the works it cites.
Mean Field Game of Controls and An Application To Trade Crowding
Cardialaguet Pierre and Lehalle Charles-Albert · 2016
Later among the works it cites.
Optimal non-asymptotic bound of the ruppert-polyak averaging without strong convexity
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Proximal splitting methods in signal processing
Patrick L Combettes and Jean-Christophe Pesquet · 2011
Cited alongside, same era.
Non-asymptotic analysis of stochastic approximation algorithms for machine learning
Eric Moulines and Francis R Bach · 2011
Cited alongside, same era.
Convergence rates of inexact proximal-gradient methods for convex optimization
Mark Schmidt, Nicolas L Roux, and Francis R Bach · 2011
Cited alongside, same era.
Adaptive algorithms and stochastic approximations
Albert Benveniste, Michel Métivier, and Pierre Priouret · 2012
Cited alongside, same era.
Markov chains and stochastic stability
Sean P Meyn and Richard L Tweedie · 2012
Cited alongside, same era.
Sébastien Gadat and Fabien Panloup · 2017
Later among the works it cites.
Limit Order Strategic Placement with Adverse Selection Risk and the Role of Latency
Charles-Albert Lehalle and Othmane Mounjid · 2017
Later among the works it cites.
Market Microstructure in Practice
Charles-Albert Lehalle, Sophie Laruelle, Romain Burgot, Stéphanie Pelin, and Matthieu Lasnier · 2018
Later among the works it cites.
DGM: A deep learning algorithm for solving partial differential equations
Justin Sirignano and Konstantinos Spiliopoulos · 2018
Later among the works it cites.
Relative deviation learning bounds and generalization with unbounded loss functions
Corinna Cortes, Spencer Greenberg, and Mehryar Mohri · 2019
Closest in time.
Deep reinforcement learning for market making in corporate bonds: beating the curse of dimensionality, June 2019
Iuliia Manziuk and Olivier Guéant · 2019
Closest in time.