Fetching the paper…
Reading the bibliography…
In this paper, we present C-ADAM, the first adaptive solver for compositional problems involving a non-linear functional nesting of expected values.
Finn, C., Rajeswaran, A., Kakade, S. M., and Levine, S · 1902
Earlier work this paper cites.
On the convergence of adam and beyond
Reddi, S. J., Kale, S., and Kumar, S · 1904
Earlier work this paper cites.
Wang, Y. and Yao, Q · 1904
Earlier work this paper cites.
A survey of optimization methods from a machine learning perspective
Sun, S., Cao, Z., Zhu, H., and Zhao, J · 1906
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
Polyak, B. T. and Juditsky, A. B · 1992
Earlier work this paper cites.
Learning to learn: Introduction and overview
Thrun, S. and Pratt, L · 1998
Earlier work this paper cites.
Convex Optimization
Boyd, S. and Vandenberghe, L · 2004
Earlier work this paper cites.
Cubic regularization of newton method and its global performance
Nesterov, Y. and Polyak, B · 2006
Earlier work this paper cites.
Sparse additive models, 2007
Ravikumar, P., Lafferty, J., Liu, H., and Wasserman, L · 2007
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, J., Hazan, E., and Singer, Y · 2011
Earlier work this paper cites.
One shot learning of simple visual concepts
Lake, B., Salakhutdinov, R., Gross, J., and Tenenbaum, J · 2011
Earlier work this paper cites.
Adadelta: An adaptive learning rate method, 2012
Zeiler, M. D · 2012
Earlier work this paper cites.
Ella: An efficient lifelong learning algorithm
Ruvolo, P. and Eaton, E · 2013
Earlier work this paper cites.
Online multi-task learning for policy gradient methods
Bou-Ammar, H., Eaton, E., Ruvolo, P., and Taylor, M. E · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization, 2014
Kingma, D. P. and Ba, J · 2014
Earlier work this paper cites.
Introductory Lectures on Convex Optimization: A Basic Course
Nesterov, Y · 2014
Earlier work this paper cites.
Stochastic compositional gradient descent: Algorithms for minimizing compositions of expected-value functions, 2014
Wang, M., Fang, E. X., and Liu, H · 2014
Earlier work this paper cites.
Autonomous cross-domain knowledge transfer in lifelong policy gradient reinforcement learning
Bou-Ammar, H., Eaton, E., Luna, J., and Ruvolo, P · 2015
Earlier work this paper cites.
Accelerated proximal gradient methods for nonconvex programming
Li, H. and Lin, Z · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
Cited alongside, same era.
Variance reduction for faster non-convex optimization, 2016
Allen-Zhu, Z. and Hazan, E · 2016
Cited alongside, same era.
End to end learning for self-driving cars
Bojarski, M., Del Testa, D., Dworakowski, D., Firner, B., Flepp, B., Goyal, P., Jackel, L. D., Monfort, M., Muller, U., Zhang, J., et al · 2016
Cited alongside, same era.
(bandit) convex optimization with biased noisy gradient oracles
Hu, X., Prashanth, L., György, A., and Szepesvári, C · 2016
Cited alongside, same era.
Finite-sum composition optimization via variance reduced gradient descent, 2016
Lian, X., Wang, M., and Liu, J · 2016
Cited alongside, same era.
Balancing two-player stochastic games with soft q-learning, 2018
Grau-Moya, J., Leibfried, F., and Bou-Ammar, H · 2018
Later among the works it cites.
Bayesian model-agnostic meta-learning
Kim, T., Yoon, J., Dia, O., Kim, S., Bengio, Y., and Ahn, S · 2018
Later among the works it cites.
Improved sample complexity for stochastic compositional variance reduced gradient, 2018
Lin, T., Fan, C., Wang, M., and Jordan, M. I · 2018
Later among the works it cites.
Stochastically controlled stochastic gradient for the convex and non-convex composition problem, 2018
Liu, L., Liu, J., Hsieh, C.-J., and Tao, D · 2018
Later among the works it cites.
A viscosity approach to stochastic differential games of control and stopping involving impulsive control, 2018
Mguni, D · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ravi, S. and Larochelle, H · 2016
Cited alongside, same era.
A distributed newton method for large scale consensus optimization
Tutunov, R., Bou-Ammar, H., and Jadbabaie, A · 2016
Cited alongside, same era.
Accelerating stochastic composition optimization
Wang, M., Liu, J., and Fang, E · 2016
Cited alongside, same era.
First-Order Methods in Optimization
Beck, A · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Accelerated method for stochastic composition optimization with nonsmooth regularization
Huo, Z., Gu, B., and Huang, H · 2017
Cited alongside, same era.
Variance reduced methods for non-convex composition optimization, 2017
Liu, L., Liu, J., and Tao, D · 2017
Cited alongside, same era.
Tian, Z., Zou, S., Warr, T., Wu, L., and Wang, J · 2018
Later among the works it cites.
Toward multimodal model-agnostic meta-learning
Vuorio, R., Sun, S.-H., Hu, H., and Lim, J. J · 2018
Later among the works it cites.
Adaptive methods for nonconvex optimization
Zaheer, M., Reddi, S., Sachan, D., Kale, S., and Kumar, S · 2018
Later among the works it cites.
Wasserstein robust reinforcement learning, 2019
Abdullah, M. A., Ren, H., Ammar, H. B., Milenkovic, V., Luo, R., Zhang, M., and Wang, J · 2019
Later among the works it cites.
Nonparametric compositional stochastic optimization, 2019
Bedi, A. S., Koppel, A., and Rajawat, K · 2019
Later among the works it cites.
Automated deep learning design for medical image classification by health-care professionals with no coding experience: a feasibility study
Faes, L., Wagner, S. K., Fu, D. J., Liu, X., Korot, E., Ledsam, J. R., Back, T., Chopra, R., Pontikos, N., Kern, C., et al · 2019
Later among the works it cites.
On the convergence theory of gradient-based model-agnostic meta-learning algorithms, 2019
Fallah, A., Mokhtari, A., and Ozdaglar, A · 2019
Later among the works it cites.
Derivative-free & order-robust optimisation
Gabillon, V., Tutunov, R., Valko, M., and Ammar, H. B · 2019
Later among the works it cites.
On the convergence of model-agnostic meta-learning
Golmant, N · 2019
Later among the works it cites.
Efficient smooth non-convex stochastic compositional optimization via stochastic recursive gradient descent
Hu, W., Li, C. J., Lian, X., Liu, J., and Yuan, H · 2019
Later among the works it cites.
A review of classical first order optimization methods in machine learning
Menghan, Z · 2019
Later among the works it cites.
Modelling bounded rationality in multi-agent interactions by generalized recursive reasoning, 2019
Wen, Y., Yang, Y., Luo, R., and Wang, J · 2019
Later among the works it cites.
Deep learning for anomaly detection
Wang, R., Nie, K., Wang, T., Yang, Y., and Long, B · 2020
Closest in time.