Fetching the paper…
Reading the bibliography…
Bilevel optimization (BO) is useful for solving a variety of important machine learning problems including but not limited to hyperparameter optimization, meta-learning, continual learning, and reinforcement learning.
Directional derivative of the marginal function in nonlinear programming
Robert Janin · 1984
Earlier work this paper cites.
On the numerical solution of a class of stackelberg problems
Jivrí V Outrata · 1990
Earlier work this paper cites.
Optimality conditions for bilevel programming problems
JJ Ye and DL Zhu · 1995
Earlier work this paper cites.
Foundations of bilevel programming
Stephan Dempe · 2002
Earlier work this paper cites.
Numerical Optimization
Jorge Nocedal and Stephen J. Wright · 2006
Earlier work this paper cites.
Subdifferentials of value functions and optimality conditions for dc and bilevel infinite and semi-infinite programs
Nguyen Dinh, B Mordukhovich, and Tran TA Nghia · 2010
Earlier work this paper cites.
The mnist database of handwritten digit images for machine learning research
Li Deng · 2012
Earlier work this paper cites.
Practical bilevel optimization: algorithms and applications , volume 30
Jonathan F Bard · 2013
Earlier work this paper cites.
Global stability of dynamical systems
Michael Shub · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Hamed Karimi, Julie Nutini, and Mark Schmidt · 2016
Earlier work this paper cites.
Gradient descent only converges to minimizers
Jason D Lee, Max Simchowitz, Michael I Jordan, and Benjamin Recht · 2016
Earlier work this paper cites.
Hyperparameter optimization with approximate gradient
Fabian Pedregosa · 2016
Earlier work this paper cites.
Constantinos Daskalakis, Andrew Ilyas, Vasilis Syrgkanis, and Haoyang Zeng · 2017
Earlier work this paper cites.
Forward and reverse gradient-based hyperparameter optimization
Luca Franceschi, Michele Donini, Paolo Frasconi, and Massimiliano Pontil · 2017
Earlier work this paper cites.
Gradient episodic memory for continual learning
David Lopez-Paz and Marc’Aurelio Ranzato · 2017
Earlier work this paper cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Earlier work this paper cites.
Bilevel programming for hyperparameter optimization and meta-learning
Luca Franceschi, Paolo Frasconi, Saverio Salzo, Riccardo Grazzi, and Massimiliano Pontil · 2018
Cited alongside, same era.
Approximation methods for bilevel programming
Saeed Ghadimi and Mengdi Wang · 2018
Cited alongside, same era.
Reviving and improving recurrent back-propagation
Renjie Liao, Yuwen Xiong, Ethan Fetaya, Lisa Zhang, KiJung Yoon, Xaq Pitkow, Raquel Urtasun, and Richard Zemel · 2018
Cited alongside, same era.
Learning to learn without forgetting by maximizing transfer and minimizing interference
Matthew Riemer, Ignacio Cases, Robert Ajemian, Miao Liu, Irina Rish, Yuhai Tu, and Gerald Tesauro · 2018
Cited alongside, same era.
On tiny episodic memories in continual learning
Arslan Chaudhry, Marcus Rohrbach, Mohamed Elhoseiny, Thalaiyasingam Ajanthan, Puneet K Dokania, Philip HS Torr, and Marc’Aurelio Ranzato · 2019
Amortized implicit differentiation for stochastic bilevel optimization
Michael Arbel and Julien Mairal · 2021
Later among the works it cites.
A single-timescale stochastic bilevel optimization method
Tianyi Chen, Yuejiao Sun, and Wotao Yin · 2021
Later among the works it cites.
Simple bilevel programming and extensions
Stephan Dempe, Nguyen Dinh, Joydeep Dutta, and Tanushree Pandit · 2021
Later among the works it cites.
Proxy convexity: A unified framework for the analysis of neural networks trained by gradient descent
Spencer Frei and Quanquan Gu · 2021
Later among the works it cites.
Tommaso Giovannelli, Griffin Kent, and Luis Nunes Vicente · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Matthew MacKay, Paul Vicol, Jon Lorraine, David Duvenaud, and Roger Grosse · 2019
Cited alongside, same era.
Meta-learning with implicit gradients
Aravind Rajeswaran, Chelsea Finn, Sham Kakade, and Sergey Levine · 2019
Cited alongside, same era.
Truncated back-propagation for bilevel optimization
Amirreza Shaban, Ching-An Cheng, Nathan Hatch, and Byron Boots · 2019
Cited alongside, same era.
Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost
Zhuoran Yang, Yongxin Chen, Mingyi Hong, and Zhaoran Wang · 2019
Cited alongside, same era.
Bilevel optimization
Stephan Dempe and Alain Zemkoho · 2020
Cited alongside, same era.
On the iteration complexity of hypergradient computation
Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil, and Saverio Salzo · 2020
Cited alongside, same era.
Mingyi Hong, Hoi-To Wai, Zhaoran Wang, and Zhuoran Yang · 2020
Cited alongside, same era.
Later among the works it cites.
Automatic and harmless regularization with constrained and lexicographic optimization: A dynamic barrier approach
Chengyue Gong, Xingchao Liu, and Qiang Liu · 2021
Later among the works it cites.
Randomized stochastic variance-reduced methods for stochastic bilevel optimization
Zhishuai Guo and Tianbao Yang · 2021
Later among the works it cites.
Lower bounds and accelerated algorithms for bilevel optimization
Kaiyi Ji and Yingbin Liang · 2021
Later among the works it cites.
Bilevel optimization: Convergence analysis and enhanced design
Kaiyi Ji, Junjie Yang, and Yingbin Liang · 2021
Later among the works it cites.
Learning to defend by learning to attack
Haoming Jiang, Zhehui Chen, Yuyang Shi, Bo Dai, and Tuo Zhao · 2021
Later among the works it cites.
A near-optimal algorithm for stochastic bilevel optimization via double-momentum
Prashant Khanduri, Siliang Zeng, Mingyi Hong, Hoi-To Wai, Zhaoran Wang, and Zhuoran Yang · 2021
Later among the works it cites.
A fully single loop algorithm for bilevel optimization without hessian inverse
Junyi Li, Bin Gu, and Heng Huang · 2021
Later among the works it cites.
Penalty method for inversion-free deep bilevel optimization
Akshay Mehra and Jihun Hamm · 2021
Later among the works it cites.
Subquadratic overparameterization for shallow neural networks
Chaehwan Song, Ali Ramezani-Kebrya, Thomas Pethick, Armin Eftekhari, and Volkan Cevher · 2021
Later among the works it cites.
Provably faster algorithms for bilevel optimization
Junjie Yang, Kaiyi Ji, and Yingbin Liang · 2021
Later among the works it cites.
Loss landscapes and optimization in over-parameterized non-linear systems and neural networks
Chaoyue Liu, Libin Zhu, and Mikhail Belkin · 2022
Closest in time.