Fetching the paper…
Reading the bibliography…
We study a class of sampled stochastic optimization problems, where the underlying state process has diffusive dynamics of the mean-field type.
1912
Earlier work this paper cites.
Hasan, M., and A.K. Roy-Chowdhury (2015): A continuous learning framework for activity recognition using deep hybrid feature models. IEEE Trans. Multimedia
1922
Earlier work this paper cites.
Srivastava, N., G.E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov (2014): Dropout: a simple way to prevent neural networks from overfitting. J. Mach. Learn. Res
1958
Earlier work this paper cites.
El Karoui, N., D.H. Nguyen, and M. Jeanblanc-Picquè (1987): Compactification methods in the control of degenerate diffusions: existence of an optimal control. Stochastics
1987
Earlier work this paper cites.
Simon, J. (1987): Compact sets in the space L p ( 0 , T , B ) L^{p}(0,T;B) . Ann. Mat. Pura Appl
1987
Earlier work this paper cites.
Haussmann, U.G., and J.P. Lepeltier (1990): On the existence of optimal controls. SIAM J. Control Optim
1990
Earlier work this paper cites.
De Figueiredo, D.G. (1991): Lectures on the Ekeland variational principle with applications and detours. Acta. Appl. Math
1991
Earlier work this paper cites.
Dal Maso, G. (1993): An Introduction to Γ \Gamma -Convergence
1993
Earlier work this paper cites.
Luçon, E., and W. Stannat (2014): Mean field limit for disordered diffusions with singular interactions. Ann. Appl. Probab
1993
Earlier work this paper cites.
Barbu, V., M. Röckner, D. Zhang (2018): Optimal bilinear control of nonlinear stochastic Schrödinger equations driven by linear multiplicative noise. Ann. Probab
1999
Earlier work this paper cites.
Villani, C. (2003): Topics in Optimal Transportation
2003
Earlier work this paper cites.
Aliprantis, C.D., and K. Border (2006): Infinite Dimensional Analysis: A Hitchhiker’s Guide
2006
Earlier work this paper cites.
Bahlali, S., B. Djehiche, and B. Mezerdi (2007): The relaxed stochastic maximum principle in singular optimal control of diffusions. SIAM J. Control Optim
2007
Earlier work this paper cites.
Haykin, S. (2009): Neural Network and Learning Machines, 3rd Ed.
2009
Earlier work this paper cites.
Villani, C. (2009): Optimal Transport: Old and New Part
2009
Earlier work this paper cites.
Annunziato, M., and A. Borzì (2010): Optimal control of probability density functions of stochastic processes. Math. Model. Anal
2010
Earlier work this paper cites.
Evans, L.C. (2010): Partial Differential Equations, 2nd Ed
2010
Earlier work this paper cites.
Ahmed, N.U., and C.D. Charalambous (2013): Stochastic minimum principle for partially observed systems subject to continuous and jump diffusion processes and driven by relaxed controls. SIAM J. Control Optim
2013
Cited alongside, same era.
Ba, J., and B. Frey (2013): Adaptive dropout for training deep neural networks. Advances in Neural Information Processing Systems (NIPS)
2013
Cited alongside, same era.
Hintermüller, M., D. Marahrens, P.A. Markowich, and C. Sparber (2013): Optimal bilinear control of Gross-Pitaevskii equations. SIAM J. Control Optim
2013
Cited alongside, same era.
2013
Cited alongside, same era.
Chizat, L. and F. Bach (2018): On the global convergence of gradient descent for over-parameterized models using optimal transport. Proceedings of 32nd Conference on Neural Information Processing Systems (NeurIPS)
2018
Later among the works it cites.
Dieng, A.B., R. Ranganath, J. Altosaar, and D.M. Blei (2018): Noisin: unbiased regularization for recurrent neural networks. Proceedings of the 35th International Conference on Machine Learning (ICML-18)
2018
Later among the works it cites.
E, Weinan, J.Q. Han, and Q.X. Li (2018): A mean-field optimal control formulation of deep learning. Res. Math. Sci
2018
Later among the works it cites.
Haber, E., and L. Ruthotto (2018): Stable architectures for deep neural networks. Inverse Problems
2018
Later among the works it cites.
Lu, Y., A. Zhong, Q. Li, and B. Dong (2018): Beyond finite layer neural networks: Bridging deep architectures and numerical differential equations. Proceedings of the 35 th International Conference on Machine Learning (PMLR)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ioffe, S., and C. Szegedy (2015): Batch normalization: accelerating deep network training by reducing internal covariate shift. ICML’15: Proceedings of the 32nd International Conference on International Conference on Machine Learning
2015
Cited alongside, same era.
Kingma, D.P., T. Salimans, and M. Welling (2015): Variational dropout and the local reparameterization trick. Advances in Neural Information Processing Systems
2015
Cited alongside, same era.
Lacker, D. (2015): Mean field games via controlled martingale problems: Existence of Markovian equilibria. Stochastic Process. Appl
2015
Cited alongside, same era.
Manita, O.A., M.S. Romanov, and S.V. Shaposhnikov (2015): On uniqueness of solutions to nonlinear Fokker-Planck-Kolmogorov equations. Nonlinear Anal
2015
Cited alongside, same era.
Carmona, R., F. Delarue, and D. Lacker (2016): Mean field games with common noise. Ann. Probab
2016
Cited alongside, same era.
He, K., S. Ren, and J. Sun (2016): Deep residual learning for image recognition. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2016
Cited alongside, same era.
Lacker, D. (2016): A general characterization of the mean field limit for stochastic differential games. Probab. Theory Related Fields
2016
Cited alongside, same era.
Mallat, S. (2016): Understanding deep convolutional neural networks. Philos. Trans. Roy. Soc. A
2016
Cited alongside, same era.
2018
Later among the works it cites.
Mei, S., A. Montanari, and P. Nguyen (2018): A mean field view of the landscape of two-layer neural networks. PNAS
2018
Later among the works it cites.
2018
Later among the works it cites.
Du, S.S., J. Lee, H. Li, L. Wang, and X. Zhai (2019): Gradient descent finds global minima of deep neural networks. Proceeding of the 36th International Conference on Machine Learning (ICML-19)
2019
Closest in time.
He, Z.H., A.S. Rakin, and D.L. Fan (2019): Parametric noise injection: Trainable randomness to improve deep neural network robustness against adversarial attack. Proceedings-2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Closest in time.
2019
Closest in time.
Cuchiero, C., M. Larsson, and J. Teichmann (2020): Deep neural networks, generic universal interpolation, and controlled ODEs. SIAM J. Math. Data Sci
2020
Closest in time.
2020
Closest in time.
Sirignano, J. and K. Spiliopoulos (2021): Mean field analysis of neural networks: A law of large numbers. SIAM J. Appl. Math
2021
Closest in time.
Bencheikh, O., and B. Jourdain (2022): Weak and strong error analysis for mean-field rank-based particle approximations of one dimensional viscous scalar conservation laws. Forthcoming in Ann. Appl. Probab
2022
Closest in time.
Motte, M., and H. Pham (2022): Mean-field Markov decision processes with common noise and open-loop controls. Ann. Appl. Probab
2022
Closest in time.
Sirignano, J., and K. Spiliopoulos (2022): Mean field analysis of deep neural networks. Math. Oper. Res
2022
Closest in time.