Fetching the paper…
Reading the bibliography…
In this paper, we consider non-convex stochastic bilevel optimization (SBO) problems that have many applications in machine learning.
Meta-learning with implicit gradients
Aravind Rajeswaran, Chelsea Finn, Sham Kakade, and Sergey Levine · 1909
Earlier work this paper cites.
Actor-critic–type learning algorithms for markov decision processes
Vijaymohan R Konda and Vivek S Borkar · 1999
Earlier work this paper cites.
Fast training of support vector machines using sequential minimal optimization
John C. Platt · 1999
Earlier work this paper cites.
An overview of bilevel optimization
Benoît Colson, Patrice Marcotte, and Gilles Savard · 2007
Earlier work this paper cites.
Classification model selection via bilevel programming
Gautam Kunapuli, Kristin P Bennett, Jing Hu, and Jong-Shi Pang · 2008
Earlier work this paper cites.
A bilevel optimization approach for parameter learning in variational models
Karl Kunisch and Thomas Pock · 2013
Earlier work this paper cites.
UCI machine learning repository, 2017
Dheeru Dua and Casey Graff · 2017
Earlier work this paper cites.
SARAH: A novel method for machine learning problems using stochastic recursive gradient
Lam M Nguyen, Jie Liu, Katya Scheinberg, and Martin Takác · 2017
Earlier work this paper cites.
A first order method for solving convex bilevel optimization problems
Shoham Sabach and Shimrit Shtern · 2017
Earlier work this paper cites.
SPIDER: near-optimal non-convex optimization via stochastic path-integrated differential estimator
Cong Fang, Chris Junchi Li, Zhouchen Lin, and Tong Zhang · 2018
Earlier work this paper cites.
Bilevel programming for hyperparameter optimization and meta-learning
Luca Franceschi, Paolo Frasconi, Saverio Salzo, Riccardo Grazzi, and Massimiliano Pontil · 2018
Cited alongside, same era.
Approximation methods for bilevel programming
Saeed Ghadimi and Mengdi Wang · 2018
Cited alongside, same era.
Accelerated method for stochastic composition optimization with nonsmooth regularization
Zhouyuan Huo, Bin Gu, Ji Liu, and Heng Huang · 2018
Cited alongside, same era.
Stochastically controlled stochastic gradient for the convex and non-convex composition problem
Liu Liu, Ji Liu, Cho-Jui Hsieh, and Dacheng Tao · 2018
Cited alongside, same era.
Lower bounds for non-convex stochastic optimization
Yossi Arjevani, Yair Carmon, John C Duchi, Dylan J Foster, Nathan Srebro, and Blake Woodworth · 2019
Solving stochastic compositional optimization is nearly as easy as solving stochastic optimization
Tianyi Chen, Yuejiao Sun, and Wotao Yin · 2020
Later among the works it cites.
On the convergence theory of gradient-based model-agnostic meta-learning algorithms
Alireza Fallah, Aryan Mokhtari, and Asuman Ozdaglar · 2020
Later among the works it cites.
Mingyi Hong, Hoi-To Wai, Zhaoran Wang, and Zhuoran Yang · 2020
Later among the works it cites.
Biased stochastic first-order methods for conditional stochastic optimization and applications in meta learning
Yifan Hu, Siqi Zhang, Xin Chen, and Niao He · 2020
Later among the works it cites.
Accelerated zeroth-order momentum methods from mini to minimax optimization
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Momentum-based variance reduction in non-convex SGD
Ashok Cutkosky and Francesco Orabona · 2019
Cited alongside, same era.
DARTS: Differentiable architecture search
Hanxiao Liu, Karen Simonyan, and Yiming Yang · 2019
Cited alongside, same era.
Truncated back-propagation for bilevel optimization
Amirreza Shaban, Ching-An Cheng, Nathan Hatch, and Byron Boots · 2019
Cited alongside, same era.
A stochastic composite gradient method with incremental variance reduction
Junyu Zhang and Lin Xiao · 2019
Cited alongside, same era.
Momentum schemes with stochastic variance reduction for nonconvex composite optimization
Yi Zhou, Zhe Wang, Kaiyi Ji, Yingbin Liang, and Vahid Tarokh · 2019
Cited alongside, same era.
A single timescale stochastic approximation method for nested stochastic optimization
Saeed Ghadimi, Andrzej Ruszczynski, and Mengdi Wang
Cited in the paper.
A single timescale stochastic approximation method for nested stochastic optimization
Saeed Ghadimi, Andrzej Ruszczynski, and Mengdi Wang
Cited in the paper.
Feihu Huang, Shangqian Gao, Jian Pei, and Heng Huang · 2020
Later among the works it cites.
Provably faster algorithms for bilevel optimization and applications to meta-learning
Kaiyi Ji, Junjie Yang, and Yingbin Liang · 2020
Later among the works it cites.
A practical online method for distributionally deep robust optimization
Qi Qi, Zhishuai Guo, Yi Xu, Rong Jin, and Tianbao Yang · 2020
Later among the works it cites.
Zhuoning Yuan, Yan Yan, Milan Sonka, and Tianbao Yang · 2020
Later among the works it cites.
A single-timescale stochastic bilevel optimization method
Tianyi Chen, Yuejiao Sun, and Wotao Yin · 2021
Closest in time.
Qi Qi, Youzhi Luo, Zhao Xu, Shuiwang Ji, and Tianbao Yang · 2021
Closest in time.