Fetching the paper…
Reading the bibliography…
Stochastic optimal control and games have a wide range of applications, from finance and economics to social sciences, robotics, and energy management.
Some notes on computation of games solutions
G. W. Brown · 1949
Earlier work this paper cites.
Iterative solution of games by fictitious play
G. W. Brown · 1951
Earlier work this paper cites.
Non-cooperative games
J. Nash · 1951
Earlier work this paper cites.
A Markovian decision process
R. Bellman · 1957
Earlier work this paper cites.
Differential Games: A Mathematical Theory with Applications to Warfare and Pursuit, Control and Optimization
R. Isaacs · 1965
Earlier work this paper cites.
Time to build and aggregate fluctuations
F. E. Kydland and E. C. Prescott · 1982
Earlier work this paper cites.
Stochastic Functional Differential Equations
S.-E. A. Mohammed · 1984
Earlier work this paper cites.
A tutorial on dynamic and differential games
T. Başar · 1986
Earlier work this paper cites.
Learning representations by back-propagating errors
D. E. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
A multilayered neural network controller
D. Psaltis, A. Sideris, and A. A. Yamamura · 1988
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
G. Cybenko · 1989
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
K. Hornik, M. Stinchcombe, and H. White · 1989
Earlier work this paper cites.
Numerical methods for stochastic control problems in continuous time
H. J. Kushner · 1990
Earlier work this paper cites.
Adapted solution of a backward stochastic differential equation
E. Pardoux and S. Peng · 1990
Earlier work this paper cites.
Approximation capabilities of multilayer feedforward networks
K. Hornik · 1991
Earlier work this paper cites.
Topics in propagation of chaos
A.-S. Sznitman · 1991
Earlier work this paper cites.
Neural networks for control systems—a survey
K. J. Hunt, D. Sbarbaro, R. Żbikowski, and P. J. Gawthrop · 1992
Earlier work this paper cites.
Stochastic Hamilton–Jacobi–Bellman equations
S. Peng · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Multilayer feedforward networks with a nonpolynomial activation function can approximate any function
M. Leshno, V. Y. Lin, A. Pinkus, and S. Schocken · 1993
Earlier work this paper cites.
Stochastic optimal control: the discrete-time case
D. P. Bertsekas and S. E. Shreve · 1996
Earlier work this paper cites.
Reinforcement learning: A survey
L. P. Kaelbling, M. L. Littman, and A. W. Moore · 1996
Earlier work this paper cites.
Control of Systems with Aftereffect
V. B. Kolmanovskiĭ and L. E. Shaĭkhet · 1996
Earlier work this paper cites.
Galerkin approximations of the generalized hamilton-jacobi-bellman equation
R. W. Beard, G. N. Saridis, and J. T. Wen · 1997
Earlier work this paper cites.
Numerical methods for backward stochastic differential equations
D. Chevance · 1997
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Piecewise affine neural networks and nonlinear control
C.-A. Lehalle and R. Azencott · 1998
Earlier work this paper cites.
Stochastic differential systems with memory: theory, examples and applications
S.-E. A. Mohammed · 1998
Earlier work this paper cites.
When are quasi-monte carlo algorithms efficient for high dimensional integrals?
I. H. Sloan and H. Woźniakowski · 1998
Earlier work this paper cites.
Time-to-build and cycles
P. K. Asea and P. J. Zak · 1999
Earlier work this paper cites.
Numerical methods for pursuit-evasion games via viscosity solutions
M. Bardi, M. Falcone, and P. Soravia · 1999
Earlier work this paper cites.
Stochastic and differential games: theory and numerical methods
M. Bardi, T. Raghavan, and T. Parthasarathy · 1999
Earlier work this paper cites.
Differential games: a mathematical theory with applications to warfare and pursuit, control and optimization
R. Isaacs · 1999
Earlier work this paper cites.
Approximation theory of the MLP model in neural networks
A. Pinkus · 1999
Earlier work this paper cites.
Relationship between backward stochastic differential equations and stochastic controls: a linear-quadratic approach
M. Kohlmann and X. Y. Zhou · 2000
Earlier work this paper cites.
A maximum principle for optimal control of stochastic systems with delay, with applications to finance
B. Øksendal and A. Sulem · 2000
Earlier work this paper cites.
A stochastic quantization method for nonlinear problems
V. Bally, J. Printems, et al · 2001
Earlier work this paper cites.
The finite element approximation of hamilton-jacobi-bellman equations
M. Boulbrachene and M. Haiour · 2001
Earlier work this paper cites.
Optimal consumption under partial observations for a stochastic system with delay
I. Elsanosi and B. Larssen · 2001
Earlier work this paper cites.
A BVP solver based on residual control and the Maltab PSE
J. Kierzenka and L. F. Shampine · 2001
Earlier work this paper cites.
Learning precise timing with LSTM recurrent networks
F. A. Gers, N. N. Schraudolph, and J. Schmidhuber · 2002
Earlier work this paper cites.
Numerical approximations for stochastic differential games
H. J. Kushner · 2002
Earlier work this paper cites.
System control and rough paths
T. Lyons and Z. Qian · 2002
Earlier work this paper cites.
Numerical method for backward stochastic differential equations
J. Ma, P. Protter, J. San Martín, and S. Torres · 2002
Earlier work this paper cites.
Error analysis of the optimal quantization algorithm for obstacle problems
V. Bally and G. Pages · 2003
Earlier work this paper cites.
Consistency of generalized finite difference schemes for the stochastic hjb equation
J. F. Bonnans and H. Zidani · 2003
Earlier work this paper cites.
Nash Q-learning for general-sum stochastic games
J. Hu and M. P. Wellman · 2003
Earlier work this paper cites.
Discrete-time approximation and Monte-Carlo simulation of backward stochastic differential equations
B. Bouchard and N. Touzi · 2004
Earlier work this paper cites.
Numerical approximations for stochastic differential games: the ergodic case
H. J. Kushner · 2004
Earlier work this paper cites.
A numerical scheme for BSDEs
J. Zhang · 2004
Earlier work this paper cites.
Stochastic control problems with delay
H. Bauer and U. Rieder · 2005
Earlier work this paper cites.
A regression-based monte carlo method to solve backward stochastic differential equations
E. Gobet, J.-P. Lemor, and X. Warin · 2005
Earlier work this paper cites.
Sensitivity analysis using Itô-Malliavin calculus and martingales, and application to stochastic optimal control
E. Gobet and R. Munos · 2005
Earlier work this paper cites.
Stochastic optimal control of delay equations arising in advertising models
F. Gozzi, S. di Roma, and C. Marinelli · 2005
Earlier work this paper cites.
On some recent aspects of stochastic control and their applications
H. Pham · 2005
Earlier work this paper cites.
Numerical methods for the pricing of swing options: a stochastic control approach
C. Barrera-Esteve, F. Bergeret, C. Dossal, E. Gobet, A. Meziou, R. Munos, and D. Reboul-Salze · 2006
Earlier work this paper cites.
Numerical methods for differential games based on partial differential equations
M. Falcone · 2006
Earlier work this paper cites.
Large population stochastic dynamic games: closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle
M. Huang, R. P. Malhamé, and P. E. Caines · 2006
Earlier work this paper cites.
Jeux à champ moyen. I. Le cas stationnaire
J.-M. Lasry and P.-L. Lions · 2006
Earlier work this paper cites.
Jeux à champ moyen. II. Horizon fini et contrôle optimal
J.-M. Lasry and P.-L. Lions · 2006
Earlier work this paper cites.
Policy gradient in continuous time
R. Munos · 2006
Earlier work this paper cites.
Convergent difference schemes for degenerate elliptic and parabolic equations: Hamilton–jacobi equations and free boundary problems
A. M. Oberman · 2006
Earlier work this paper cites.
Recurrent neural networks are universal approximators
A. M. Schäfer and H. G. Zimmermann · 2006
Earlier work this paper cites.
A new kind of accurate numerical method for backward stochastic differential equations
W. Zhao, L. Chen, and S. Peng · 2006
Earlier work this paper cites.
Convergent numerical scheme for singular stochastic control with state constraints in a portfolio selection problem
A. Budhiraja and K. Ross · 2007
Earlier work this paper cites.
Second-order backward stochastic differential equations and fully nonlinear parabolic PDEs
P. Cheridito, H. M. Soner, N. Touzi, and N. Victoir · 2007
Earlier work this paper cites.
Numerical methods for controlled Hamilton-Jacobi-Bellman PDEs in finance
P. A. Forsyth and G. Labahn · 2007
Earlier work this paper cites.
Large-population cost-coupled LQG problems with nonuniform agents: individual-mass behavior and decentralized ϵ \epsilon -Nash equilibria
M. Huang, P. E. Caines, and R. P. Malhamé · 2007
Earlier work this paper cites.
Numerical approximations for nonzero-sum stochastic differential games
H. J. Kushner · 2007
Earlier work this paper cites.
Mean field games
J.-M. Lasry and P.-L. Lions · 2007
Earlier work this paper cites.
Mean field games
J.-M. Lasry and P.-L. Lions · 2007
Earlier work this paper cites.
Differential equations driven by rough paths
T. J. Lyons, M. Caruana, and T. Lévy · 2007
Earlier work this paper cites.
Approximate Dynamic Programming: Solving the curses of dimensionality
W. B. Powell · 2007
Earlier work this paper cites.
A comprehensive survey of multiagent reinforcement learning
L. Busoniu, R. Babuska, and B. De Schutter · 2008
Earlier work this paper cites.
A semi-Lagrangian approach for natural gas storage valuation and optimal operation
Z. Chen and P. A. Forsyth · 2008
Earlier work this paper cites.
Discrete-time approximation of bsdes and probabilistic schemes for fully nonlinear pdes
B. Bouchard, R. Elie, and N. Touzi · 2009
Earlier work this paper cites.
The complexity of computing a Nash equilibrium
C. Daskalakis, P. W. Goldberg, and C. H. Papadimitriou · 2009
Earlier work this paper cites.
On controlled linear diffusions with delay in a model of optimal advertising under uncertainty with memory effects
F. Gozzi, C. Marinelli, and S. Savin · 2009
Earlier work this paper cites.
Offline handwriting recognition with multidimensional recurrent neural networks
A. Graves and J. Schmidhuber · 2009
Earlier work this paper cites.
Continuous-time stochastic control and optimization with financial applications
H. Pham · 2009
Earlier work this paper cites.
Mean field games: numerical methods
Y. Achdou and I. Capuzzo-Dolcetta · 2010
Earlier work this paper cites.
Mean field games: numerical methods
Y. Achdou and I. Capuzzo-Dolcetta · 2010
Earlier work this paper cites.
Error propagation for approximate policy and value iteration
A.-m. Farahmand, C. Szepesvári, and R. Munos · 2010
Earlier work this paper cites.
Pricing of high-dimensional American options by neural networks
M. Kohler, A. Krzyżak, and N. Todorovic · 2010
Earlier work this paper cites.
A stable multistep scheme for solving backward stochastic differential equations
W. Zhao, G. Zhang, and L. Ju · 2010
Earlier work this paper cites.
A maximum principle for SDEs of mean-field type
D. Andersson and B. Djehiche · 2011
Earlier work this paper cites.
A stochastic control problem with delay arising in a pension fund model
S. Federico · 2011
Earlier work this paper cites.
Cours du Collège de France
P.-L. Lions · 2011
Earlier work this paper cites.
Mean field games: numerical methods for the planning problem
Y. Achdou, F. Camilli, and I. Capuzzo-Dolcetta · 2012
Earlier work this paper cites.
Least-squares monte carlo for backward sdes
C. Bender and J. Steiner · 2012
Earlier work this paper cites.
T. Degris, M. White, and R. S. Sutton · 2012
Earlier work this paper cites.
Mean field for Markov decision processes: from discrete to continuous optimization
N. Gast, B. Gaujal, and J.-Y. Le Boudec · 2012
Earlier work this paper cites.
Multiagent learning: Basics, challenges, and prospects
K. Tuyls and G. Weiss · 2012
Earlier work this paper cites.
Mean field games: convergence of a finite difference method
Y. Achdou, F. Camilli, and I. Capuzzo-Dolcetta · 2013
Earlier work this paper cites.
A posteriori estimates for backward SDEs
C. Bender and J. Steiner · 2013
Earlier work this paper cites.
Mean field games and mean field type control theory
A. Bensoussan, J. Frehse, and S. C. P. Yam · 2013
Earlier work this paper cites.
Notes on mean field games, 2013
P. Cardaliaguet · 2013
Earlier work this paper cites.
Probabilistic analysis of mean-field games
R. Carmona and F. Delarue · 2013
Earlier work this paper cites.
Control of McKean-Vlasov dynamics versus mean field games
R. Carmona, F. Delarue, and A. Lachapelle · 2013
Earlier work this paper cites.
Contract theory in continuous-time models
J. Cvitanić and J. Zhang · 2013
Earlier work this paper cites.
Semi-lagrangian schemes for linear and fully non-linear diffusion equations
K. Debrabant and E. R. Jakobsen · 2013
Earlier work this paper cites.
Recent developments in numerical methods for fully nonlinear second order partial differential equations
X. Feng, R. Glowinski, and M. Neilan · 2013
Earlier work this paper cites.
Generating sequences with recurrent neural networks
A. Graves · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
A. Graves, A.-r. Mohamed, and G. Hinton · 2013
Earlier work this paper cites.
On the convergence of finite element methods for hamilton–jacobi–bellman equations
M. Jensen and I. Smears · 2013
Earlier work this paper cites.
Numerical Methods for Stochastic Control Problems in Continuous Time
H. Kushner and P. Dupuis · 2013
Cited alongside, same era.
State estimation and control of electric loads to manage real-time energy imbalance
J. L. Mathieu, S. Koch, and D. S. Callaway · 2013
Cited alongside, same era.
A fully discrete semi-Lagrangian scheme for a first order mean field game problem
E. Carlini and F. J. Silva · 2014
Cited alongside, same era.
A fully discrete semi-Lagrangian scheme for a first order mean field game problem
E. Carlini and F. J. Silva · 2014
Cited alongside, same era.
The master equation for large population equilibriums
R. Carmona and F. Delarue · 2014
Cited alongside, same era.
Q-learning in regularized mean-field games
B. Anahtarci, C. D. Kariksiz, and N. Saldi · 2020
Later among the works it cites.
Unified reinforcement Q-learning for mean field game and control problems
A. Angiuli, J.-P. Fouque, and M. Laurière · 2020
Later among the works it cites.
Unified reinforcement Q-learning for mean field game and control problems
A. Angiuli, J.-P. Fouque, and M. Laurière · 2020
Later among the works it cites.
Pricing and hedging American-style options with deep learning
S. Becker, P. Cheridito, and A. Jentzen · 2020
Later among the works it cites.
Connecting GANs, mean-field games, and optimal transport
H. Cao, X. Guo, and M. Laurière · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J.-F. Chassagneux, D. Crisan, and F. Delarue · 2014
Cited alongside, same era.
Linear multistep schemes for bsdes
J.-F. Chassagneux · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Cited alongside, same era.
On the existence of classical solutions for stationary extended mean field games
D. A. Gomes, S. Patrizi, and V. Voskanyan · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Collective target tracking mean field control for Markovian jump-driven models of electric water heating loads
A. C. Kizilkale and R. P. Malhame · 2014
Cited alongside, same era.
Dynamic programming for mean-field type control
M. Laurière and O. Pironneau · 2014
Cited alongside, same era.
Later among the works it cites.
Deep learning for constrained utility maximisation
A. Davey and H. Zheng · 2020
Later among the works it cites.
On the convergence of model free learning in mean field games
R. Elie, J. Pérolat, M. Laurière, M. Geist, and O. Pietquin · 2020
Later among the works it cites.
Deep learning methods for mean field control problems with delay
J.-P. Fouque and Z. Zhang · 2020
Later among the works it cites.
Deep fictitious play for finding Markovian Nash equilibrium in multi-agent games
J. Han and R. Hu · 2020
Later among the works it cites.
Convergence of the deep BSDE method for coupled FBSDEs
J. Han and J. Long · 2020
Later among the works it cites.
Deep learning for ranking response surfaces with applications to optimal stopping problems
R. Hu · 2020
Later among the works it cites.
Deep backward schemes for high-dimensional nonlinear PDEs
C. Huré, H. Pham, and X. Warin · 2020
Later among the works it cites.
A proof that rectified deep neural networks overcome the curse of dimensionality in the numerical approximation of semilinear heat equations
M. Hutzenthaler, A. Jentzen, T. Kruse, and T. A. Nguyen · 2020
Later among the works it cites.
Three algorithms for solving high-dimensional fully coupled FBSDEs through deep learning
S. Ji, S. Peng, Y. Peng, and X. Zhang · 2020
Later among the works it cites.
Many-player games of optimal consumption and investment under relative performance criteria
D. Lacker and A. Soret · 2020
Later among the works it cites.
On numerical methods for mean field games and mean field type control
M. Laurière · 2020
Later among the works it cites.
Learning a functional control for high-frequency finance
L. Leal, M. Laurière, and C.-A. Lehalle · 2020
Later among the works it cites.
A. T. Lin, S. W. Fung, W. Li, L. Nurbekyan, and S. J. Osher · 2020
Later among the works it cites.
Conditional optimal stopping: a time-inconsistent optimization
M. Nutz and Y. Zhang · 2020
Later among the works it cites.
Minimum width for universal approximation
S. Park, C. Yun, J. Lee, and J. Shin · 2020
Later among the works it cites.
A machine learning framework for solving high-dimensional mean field game and mean field control problems
L. Ruthotto, S. J. Osher, W. Li, L. Nurbekyan, and S. W. Fung · 2020
Later among the works it cites.
Reinforcement learning in continuous time and space: A stochastic control approach
H. Wang, T. Zariphopoulou, and X. Y. Zhou · 2020
Later among the works it cites.
Continuous-time mean–variance portfolio selection: A reinforcement learning framework
H. Wang and X. Y. Zhou · 2020
Later among the works it cites.
An overview of multi-agent reinforcement learning from game theoretical perspective
Y. Yang and J. Wang · 2020
Later among the works it cites.
Universality of deep convolutional neural networks
D.-X. Zhou · 2020
Later among the works it cites.
Deep neural networks algorithms for stochastic control problems on finite horizon: numerical applications
A. Bachouch, C. Huré, N. Langrené, and H. Pham · 2021
Later among the works it cites.
Approximation of an optimal control problem for the time-fractional fokker-planck equation
F. Camilli, S. Duisembay, and Q. Tang · 2021
Later among the works it cites.
Optimal execution with quadratic variation inventories
R. Carmona and L. Leal · 2021
Later among the works it cites.
Large-scale multi-agent deep FBSDEs
T. Chen, Z. O. Wang, I. Exarchos, and E. Theodorou · 2021
Later among the works it cites.
Approximately solving mean field games via entropy-regularized deep reinforcement learning
K. Cui and H. Koeppl · 2021
Later among the works it cites.
Exploration noise for learning linear-quadratic mean field games
F. Delarue and A. Vasileiadis · 2021
Later among the works it cites.
Neural networks-based algorithms for stochastic control and PDEs in finance
M. Germain, H. Pham, and X. Warin · 2021
Later among the works it cites.
P. Grohs, S. Ibragimov, A. Jentzen, and S. Koppensteiner · 2021
Later among the works it cites.
Multi-agent deep reinforcement learning: a survey
S. Gronauer and K. Diepold · 2021
Later among the works it cites.
Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis
H. Gu, X. Guo, X. Wei, and R. Xu · 2021
Later among the works it cites.
Mean-field multi-agent reinforcement learning: A decentralized network approach
H. Gu, X. Guo, X. Wei, and R. Xu · 2021
Later among the works it cites.
Policy gradient methods for the noisy linear quadratic regulator over a finite horizon
B. Hambly, R. Xu, and H. Yang · 2021
Later among the works it cites.
Recent advances in reinforcement learning in finance
B. Hambly, R. Xu, and H. Yang · 2021
Later among the works it cites.
Recurrent neural networks for stochastic control problems with delay
J. Han and R. Hu · 2021
Later among the works it cites.
DeepHAM: A global solution method for heterogeneous agent models with aggregate shocks
J. Han, Y. Yang, and W. E · 2021
Later among the works it cites.
Deep fictitious play for stochastic differential games
R. Hu · 2021
Later among the works it cites.
Deep neural networks algorithms for stochastic control problems on finite horizon: convergence analysis
C. Huré, H. Pham, A. Bachouch, and N. Langrené · 2021
Later among the works it cites.
A proof that deep artificial neural networks overcome the curse of dimensionality in the numerical approximation of Kolmogorov partial differential equations with constant diffusion and nonlinear drift coefficients
A. Jentzen, D. Salimova, and T. Welti · 2021
Later among the works it cites.
Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach
Y. Jia and X. Y. Zhou · 2021
Later among the works it cites.
Policy gradient and actor-critic learning in continuous time and space: Theory and algorithms
Y. Jia and X. Y. Zhou · 2021
Later among the works it cites.
On the universality of graph neural networks on large random graphs
N. Keriven, A. Bietti, and S. Vaiter · 2021
Later among the works it cites.
Neural network regression for Bermudan option pricing
B. Lapeyre and J. Lelong · 2021
Later among the works it cites.
Linear-quadratic stochastic delayed control and deep learning resolution
W. Lefebvre and E. Miller · 2021
Later among the works it cites.
Signatured deep fictitious play for mean field games with common noise
M. Min and R. Hu · 2021
Later among the works it cites.
Generalization in mean field games by learning master policies
S. Perrin, M. Laurière, J. Pérolat, R. Élie, M. Geist, and O. Pietquin · 2021
Later among the works it cites.
Mean field games flock! The reinforcement learning way
S. Perrin, M. Laurière, J. Pérolat, M. Geist, R. Élie, and O. Pietquin · 2021
Later among the works it cites.
Neural networks-based backward scheme for fully nonlinear PDEs
H. Pham, X. Warin, and M. Germain · 2021
Later among the works it cites.
A fast iterative PDE-based algorithm for feedback controls of nonsmooth mean-field control problems
C. Reisinger, W. Stockinger, and Y. Zhang · 2021
Later among the works it cites.
Path-dependent deep Galerkin method: A neural network approach to solve path-dependent partial differential equations
Y. F. Saporito and Z. Zhang · 2021
Later among the works it cites.
Learning while playing in mean-field games: Convergence and optimality
Q. Xie, Z. Yang, Z. Wang, and A. Minca · 2021
Later among the works it cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
K. Zhang, Z. Yang, and T. Başar · 2021
Later among the works it cites.
Actor-critic method for high dimensional static hamilton–jacobi–bellman partial differential equations based on neural networks
M. Zhou, J. Han, and J. Lu · 2021
Later among the works it cites.
Extensions of the deep galerkin method
A. Al-Aradi, A. Correia, G. Jardim, D. de Freitas Naiff, and Y. Saporito · 2022
Later among the works it cites.
Finite state graphon games with applications to epidemics
A. Aurell, R. Carmona, G. Dayanıklı, and M. Laurière · 2022
Later among the works it cites.
Optimal incentives to mitigate epidemics: a stackelberg mean field game approach
A. Aurell, R. Carmona, G. Dayanikli, and M. Lauriere · 2022
Later among the works it cites.
Finite approximations and Q learning for Mean Field Type Multi Agent Control
E. Bayraktar, N. Bauerle, and A. D. Kara · 2022
Later among the works it cites.
Deep signature algorithm for path-dependent american option pricing
E. Bayraktar, Q. Feng, and Z. Zhang · 2022
Later among the works it cites.
The Modern Mathematics of Deep Learning
J. Berner, P. Grohs, G. Kutyniok, and P. Petersen · 2022
Later among the works it cites.
Deep learning for mean field games and mean field control with applications to finance
R. Carmona and M. Laurière · 2022
Later among the works it cites.
Fast global convergence of natural policy gradient methods with entropy regularization
S. Cen, C. Cheng, Y. Chen, Y. Wei, and Y. Chi · 2022
Later among the works it cites.
Deep runge-kutta schemes for bsdes
J.-F. Chassagneux, J. Chen, and N. Frikha · 2022
Later among the works it cites.
Dynamics of market making algorithms in dealer markets: Learning and tacit collusion
R. Cont and W. Xiong · 2022
Later among the works it cites.
Error estimates for physics informed neural networks approximating the navier-stokes equations
T. De Ryck, A. D. Jagtap, and S. Mishra · 2022
Later among the works it cites.
Error analysis for physics-informed neural networks (PINNs) approximating Kolmogorov PDEs
T. De Ryck and S. Mishra · 2022
Later among the works it cites.
Generic bounds on the approximation error for physics-informed (and) operator learning
T. De Ryck and S. Mishra · 2022
Later among the works it cites.
Exploratory LQG mean field games with entropy regularization
D. Firoozi and S. Jaimungal · 2022
Later among the works it cites.
Convergence of the backward deep BSDE method with applications to optimal stopping problems
C. Gao, S. Gao, R. Hu, and Z. Zhu · 2022
Later among the works it cites.
DeepSets and their derivative networks for solving symmetric PDEs
M. Germain, M. Laurière, H. Pham, and X. Warin · 2022
Later among the works it cites.
Numerical resolution of McKean-Vlasov FBSDEs using neural networks
M. Germain, J. Mikael, and X. Warin · 2022
Later among the works it cites.
Machine learning architectures for price formation models
D. Gomes, J. Gutiérrez, and M. Laurière · 2022
Later among the works it cites.
Entropy regularization for mean field games with learning
X. Guo, R. Xu, and T. Zariphopoulou · 2022
Later among the works it cites.
Convergence of deep fictitious play for stochastic differential games
J. Han, R. Hu, and J. Long · 2022
Later among the works it cites.
N N -player and mean-field games in Itô-diffusion markets with competitive or homophilous interaction
R. Hu and T. Zariphopoulou · 2022
Later among the works it cites.
Robust risk-aware reinforcement learning
S. Jaimungal, S. M. Pesenti, Y. S. Wang, and H. Tatsat · 2022
Later among the works it cites.
A survey of numerical solutions for stochastic control problems: Some recent progress
Z. Jin, M. Qiu, K. Q. Tran, and G. Yin · 2022
Later among the works it cites.
On classical solutions to the mean field game system of controls
Z. Kobeissi · 2022
Later among the works it cites.
Z. Kobeissi and F. Bach · 2022
Later among the works it cites.
Learning mean field games: A survey
M. Laurière, S. Perrin, M. Geist, and O. Pietquin · 2022
Later among the works it cites.
Scalable deep reinforcement learning algorithms for mean field games
M. Laurière, S. Perrin, S. Girgin, P. Muller, A. Jain, T. Cabannes, G. Piliouras, J. Pérolat, R. Élie, O. Pietquin, et al · 2022
Later among the works it cites.
Convergence of large population games to mean field games with interaction through the controls
M. Lauriere and L. Tangpi · 2022
Later among the works it cites.
Estimates on the generalization error of physics-informed neural networks for approximating PDEs
S. Mishra and R. Molinaro · 2022
Later among the works it cites.
Estimates on the generalization error of physics-informed neural networks for approximating a class of inverse problems for PDEs
S. Mishra and R. Molinaro · 2022
Later among the works it cites.
Deep stochastic optimization in finance
A. M. Reppen, H. M. Soner, and V. Tissot-Daguette · 2022
Later among the works it cites.
Neural optimal stopping boundary
A. M. Reppen, H. M. Soner, and V. Tissot-Daguette · 2022
Later among the works it cites.
Optimally weighted loss functions for solving PDEs with neural networks
R. van der Meer, C. W. Oosterlee, and A. Borovykh · 2022
Later among the works it cites.
Optimal policies for a pandemic: A stochastic game approach and a deep learning algorithm
Y. Xuan, R. Balkin, J. Han, R. Hu, and H. D. Ceniceros · 2022
Later among the works it cites.
Pandemic control, game theory and machine learning
Y. Xuan, R. Balkin, J. Han, R. Hu, and H. D. Ceniceros · 2022
Later among the works it cites.
A machine learning enhanced algorithm for the optimal landing problem
Y. Zang, J. Long, X. Zhang, W. Hu, E. Weinan, and J. Han · 2022
Later among the works it cites.
Deep learning for Mean Field Games with non-separable Hamiltonians
M. Assouli and B. Missaoui · 2023
Closest in time.
A deep learning analysis of climate change, innovation, and uncertainty
M. Barnett, W. Brock, L. P. Hansen, R. Hu, and J. Huang · 2023
Closest in time.
Model-free mean-field reinforcement learning: mean-field mdp and mean-field q-learning
R. Carmona, M. Laurière, and Z. Tan · 2023
Closest in time.
Numerical methods for backward stochastic differential equations: A survey
J. Chessari, R. Kawai, Y. Shinozaki, and T. Yamada · 2023
Closest in time.
A class of dimensionality-free metrics for the convergence of empirical measures
J. Han, R. Hu, and J. Long · 2023
Closest in time.
Learning high-dimensional McKean-Vlasov forward-backward stochastic differential equations with general distribution dependence
J. Han, R. Hu, and J. Long · 2023
Closest in time.
Deep neural network solution for finite state mean field game with error estimation
J. Luo and H. Zheng · 2023
Closest in time.
Directed chain generative adversarial networks
M. Min, R. Hu, and T. Ichiba · 2023
Closest in time.
A posteriori error estimates for fully coupled McKean-Vlasov forward-backward SDEs
C. Reisinger, W. Stockinger, and Y. Zhang · 2023
Closest in time.
M. Zhou and J. Lu · 2023
Closest in time.
Machine learning for continuous-time finance
V. Duarte, D. Duarte, and D. Silva · 2024
Closest in time.
M. Zhou and J. Lu · 2024
Closest in time.