Fetching the paper…
Reading the bibliography…
Due to the hierarchical structure of many machine learning problems, bilevel programming is becoming more and more important recently, however, the complicated correlation between the inner and outer problem makes it extremely challenging to solve.
Solutions of ill-posed problems (an tikhonov and vy arsenin)
R. A. Willoughby · 1979
Earlier work this paper cites.
Finite perturbation of convex programs
M. C. Ferris and O. L. Mangasarian · 1991
Earlier work this paper cites.
Viscosity approximation methods for nonexpansive mappings
H.-K. Xu · 2004
Earlier work this paper cites.
Well-posed optimization problems
A. L. Dontchev and T. Zolezzi · 2006
Earlier work this paper cites.
An explicit descent method for bilevel convex optimization
M. Solodov · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky, G. Hinton, et al · 2009
Earlier work this paper cites.
Mnist handwritten digit database
Y. LeCun, C. Cortes, and C. Burges · 2010
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
J. S. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng · 2011
Earlier work this paper cites.
Minimizing the moreau envelope of nonsmooth convex functions over the fixed point set of certain quasi-nonexpansive mappings
I. Yamada, M. Yukawa, and M. Yamagishi · 2011
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Topology , volume 1
K. Kuratowski · 2014
Earlier work this paper cites.
Sequential model-based ensemble optimization
A. Lacoste, H. Larochelle, F. Laviolette, and M. Marchand · 2014
Earlier work this paper cites.
Siamese neural networks for one-shot image recognition
G. Koch, R. Zemel, and R. Salakhutdinov · 2015
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
B. M. Lake, R. Salakhutdinov, and J. B. Tenenbaum · 2015
Cited alongside, same era.
Gradient-based hyperparameter optimization through reversible learning
D. Maclaurin, D. Duvenaud, and R. Adams · 2015
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
A. Radford, L. Metz, and S. Chintala · 2015
Cited alongside, same era.
On the convergence of stochastic bi-level gradient methods
N. Couellan and W. Wang · 2016
Cited alongside, same era.
Scalable gradient-based tuning of continuous regularization hyperparameters
Improved training of wasserstein gans
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. C. Courville · 2017
Later among the works it cites.
A simple neural attentive meta-learner
N. Mishra, M. Rohaninejad, X. Chen, and P. Abbeel · 2017
Later among the works it cites.
A first order method for solving convex bilevel optimization problems
S. Sabach and S. Shtern · 2017
Later among the works it cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
H. Xiao, K. Rasul, and R. Vollgraf · 2017
Later among the works it cites.
Large scale gan training for high fidelity natural image synthesis
A. Brock, J. Donahue, and K. Simonyan · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Luketina, M. Berglund, K. Greff, and T. Raiko · 2016
Cited alongside, same era.
Hyperparameter optimization with approximate gradient
F. Pedregosa · 2016
Cited alongside, same era.
Meta-learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap · 2016
Cited alongside, same era.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra, et al · 2016
Cited alongside, same era.
M. Arjovsky, S. Chintala, and L. Bottou · 2017
Cited alongside, same era.
Online learning rate adaptation with hypergradient descent
A. G. Baydin, R. Cornish, D. M. Rubio, M. Schmidt, and F. Wood · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
C. Finn, P. Abbeel, and S. Levine · 2017
Cited alongside, same era.
Bilevel programming for hyperparameter optimization and meta-learning
L. Franceschi, P. Frasconi, S. Salzo, R. Grazzi, and M. Pontil · 2018
Later among the works it cites.
Approximation methods for bilevel programming
S. Ghadimi and M. Wang · 2018
Later among the works it cites.
Stochastic hyperparameter optimization through hypernetworks
J. Lorraine and D. Duvenaud · 2018
Later among the works it cites.
Truncated back-propagation for bilevel optimization
A. Shaban, C.-A. Cheng, N. Hatch, and B. Boots · 2018
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Later among the works it cites.
Cold case: The lost mnist digits
C. Yadav and L. Bottou · 2019
Later among the works it cites.
A generic first-order algorithmic framework for bi-level programming beyond lower-level singleton
R. Liu, P. Mu, X. Yuan, S. Zeng, and J. Zhang · 2020
Closest in time.