Fetching the paper…
Reading the bibliography…
Bilevel optimization has recently regained interest owing to its applications in emerging machine learning fields such as hyperparameter optimization, meta-learning, and reinforcement learning.
Mathematical programs with optimization problems in the constraints
Jerome Bracken and James T McGill · 1973
Earlier work this paper cites.
On the perturbation of pseudo-inverses, projections and linear least squares problems
Gilbert W Stewart · 1977
Earlier work this paper cites.
The generalised inverse
M James · 1978
Earlier work this paper cites.
Directional derivative of the marginal function in nonlinear programming
Robert Janin · 1984
Earlier work this paper cites.
The polynomial hierarchy and a simple model for competitive analysis
Robert G Jeroslow · 1985
Earlier work this paper cites.
Optimization and nonsmooth analysis
Frank H Clarke · 1990
Earlier work this paper cites.
Fast exact multiplication by the hessian
Barak A Pearlmutter · 1994
Earlier work this paper cites.
Optimality conditions for bilevel programming problems
Jane J Ye and Daoli Zhu · 1995
Earlier work this paper cites.
Mathematical Programs with Equilibrium Constraints
Zhi-Quan Luo, Jong-Shi Pang, and Daniel Ralph · 1996
Earlier work this paper cites.
Foundations of bilevel programming
Stephan Dempe · 2002
Earlier work this paper cites.
Optimality conditions for bilevel programming problems
Stephan Dempe, Vyatcheslav V Kalashnikov, and Nataliya Kalashnykova · 2006
Earlier work this paper cites.
Bilevel programming with convex lower level problems
Joydeep Dutta and Stephan Dempe · 2006
Earlier work this paper cites.
New necessary optimality conditions for bilevel programs by combining the MPEC and value function approaches
Jane J Ye and Daoli Zhu · 2010
Earlier work this paper cites.
On the karush–kuhn–tucker reformulation of the bilevel optimization problem
Stephan Dempe and Alain B Zemkoho · 2012
Earlier work this paper cites.
The bilevel programming problem: reformulations, constraint qualifications and optimality conditions
Stephan Dempe and Alain B Zemkoho · 2013
Earlier work this paper cites.
On solving simple bilevel programs with a nonconvex lower level program
Gui-Hua Lin, Mengwei Xu, and Jane J Ye · 2014
Earlier work this paper cites.
Gradient-based hyperparameter optimization through reversible learning
Dougal Maclaurin, David Duvenaud, and Ryan Adams · 2015
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Hamed Karimi, Julie Nutini, and Mark Schmidt · 2016
Earlier work this paper cites.
Hyperparameter optimization with approximate gradient
Fabian Pedregosa · 2016
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Forward and reverse gradient-based hyperparameter optimization
Luca Franceschi, Michele Donini, Paolo Frasconi, and Massimiliano Pontil · 2017
Earlier work this paper cites.
A first order method for solving convex bilevel optimization problems
Shoham Sabach and Shimrit Shtern · 2017
Cited alongside, same era.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham Kakade, and Mehran Mesbahi · 2018
Cited alongside, same era.
Bilevel programming for hyperparameter optimization and meta-learning
Luca Franceschi, Paolo Frasconi, Saverio Salzo, Riccardo Grazzi, and Massimilano Pontil · 2018
Cited alongside, same era.
Approximation methods for bilevel programming
Saeed Ghadimi and Mengdi Wang · 2018
Cited alongside, same era.
A geometric analysis of phase retrieval
Ju Sun, Qing Qu, and John Wright · 2018
Cited alongside, same era.
Reinforcement Learning: An Introduction
Richard S Sutton and Andrew G Barto · 2018
Provably faster algorithms for bilevel optimization
Junjie Yang, Kaiyi Ji, and Yingbin Liang · 2021
Later among the works it cites.
A single-timescale method for stochastic bilevel optimization
Tianyi Chen, Yuejiao Sun, Quan Xiao, and Wotao Yin · 2022
Later among the works it cites.
Bilevel methods for image reconstruction
Caroline Crockett and Jeffrey Fessler · 2022
Later among the works it cites.
A framework for bilevel optimization that enables stochastic and global variance reduction algorithms
Mathieu Dagréou, Pierre Ablin, Samuel Vaiter, and Thomas Moreau · 2022
Later among the works it cites.
Value function based difference-of-convex algorithm for bilevel hyperparameter selection problems
Lucy L Gao, Jane Ye, Haian Yin, Shangzhi Zeng, and Jin Zhang · 2022
Later among the works it cites.
A fully single loop algorithm for bilevel optimization without hessian inverse
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
DARTS: Differentiable architecture search
Hanxiao Liu, Karen Simonyan, and Yiming Yang · 2019
Cited alongside, same era.
Solving a class of non-convex min-max games using iterative first order methods
Maher Nouiehed, Maziar Sanjabi, Tianjian Huang, Jason D Lee, and Meisam Razaviyayn · 2019
Cited alongside, same era.
Overparameterized nonlinear learning: Gradient descent takes the shortest path?
Samet Oymak and Mahdi Soltanolkotabi · 2019
Cited alongside, same era.
Provable representation learning for imitation learning via bi-level optimization
Sanjeev Arora, Simon Du, Sham Kakade, Yuping Luo, and Nikunj Saunshi · 2020
Cited alongside, same era.
Coresets via bilevel optimization for continual learning and streaming
Zalán Borsos, Mojmír Mutnỳ, and Andreas Krause · 2020
Cited alongside, same era.
Lower bounds for finding stationary points i
Yair Carmon, John C Duchi, Oliver Hinder, and Aaron Sidford · 2020
Cited alongside, same era.
Junyi Li, Bin Gu, and Heng Huang · 2022
Later among the works it cites.
A stochastic linearized augmented lagrangian method for decentralized bilevel optimization
Songtao Lu, Siliang Zeng, Xiaodong Cui, Mark Squillante, Lior Horesh, Brian Kingsbury, Jia Liu, and Mingyi Hong · 2022
Later among the works it cites.
A constrained optimization approach to bilevel optimization with multiple inner minima
Daouda Sow, Kaiyi Ji, Ziwei Guan, and Yingbin Liang · 2022
Later among the works it cites.
FEDNEST: Federated bilevel, minimax, and compositional optimization
Davoud Ataee Tarzanagh, Mingchen Li, Christos Thrampoulidis, and Samet Oymak · 2022
Later among the works it cites.
An implicit gradient-type method for linearly constrained bilevel problems
Ioannis Tsaknakis, Prashant Khanduri, and Mingyi Hong · 2022
Later among the works it cites.
On implicit bias in overparameterized bilevel optimization
Paul Vicol, Jonathan P Lorraine, Fabian Pedregosa, David Duvenaud, and Roger B Grosse · 2022
Later among the works it cites.
Decentralized gossip-based stochastic bilevel optimization over communication networks
Shuoguang Yang, Xuezhou Zhang, and Mengdi Wang · 2022
Later among the works it cites.
Difference of convex algorithms for bilevel programs with applications in hyperparameter selection
Jane J Ye, Xiaoming Yuan, Shangzhi Zeng, and Jin Zhang · 2022
Later among the works it cites.
Revisiting and advancing fast adversarial training through the lens of bi-level optimization
Yihua Zhang, Guanhua Zhang, Prashant Khanduri, Mingyi Hong, Shiyu Chang, and Sijia Liu · 2022
Later among the works it cites.
Nyström method for accurate and scalable implicit differentiation
Ryuichiro Hataya and Makoto Yamada · 2023
Closest in time.
A two-timescale stochastic algorithm framework for bilevel optimization: Complexity analysis and application to actor-critic
Mingyi Hong, Hoi-To Wai, Zhaoran Wang, and Zhuoran Yang · 2023
Closest in time.
On momentum-based gradient methods for bilevel optimization with nonconvex lower-level
Feihu Huang · 2023
Closest in time.
Averaged method of multipliers for bi-level optimization without lower-level strong convexity
Risheng Liu, Yaohua Liu, Wei Yao, Shangzhi Zeng, and Jin Zhang · 2023
Closest in time.
First-order penalty methods for bilevel optimization
Zhaosong Lu and Sanyou Mei · 2023
Closest in time.
On penalty-based bilevel gradient descent method
Han Shen, Quan Xiao, and Tianyi Chen · 2023
Closest in time.
Alternating implicit projected sgd and its efficient variants for equality-constrained bilevel optimization
Quan Xiao, Han Shen, Wotao Yin, and Tianyi Chen · 2023
Closest in time.