Fetching the paper…
Reading the bibliography…
Gradient-based multilevel optimization (MLO) has gained attention as a framework for studying numerous problems, ranging from hyperparameter optimization and meta-learning to neural architecture search and reinforcement learning.
Torchmeta: A Meta-Learning library for PyTorch, 2019
Tristan Deleu, Tobias Würfl, Mandana Samiei, Joseph Paul Cohen, and Yoshua Bengio · 1909
Earlier work this paper cites.
Meta label correction for learning with weak supervision
Guoqing Zheng, Ahmed Hassan Awadallah, and Susan T. Dumais · 1911
Earlier work this paper cites.
Bilevel and multilevel programming: A bibliography review
Luis N Vicente and Paul H Calamai · 1994
Earlier work this paper cites.
Multilevel optimization: algorithms and applications , volume 20
Athanasios Migdalas, Panos M Pardalos, and Peter Värbrand · 1998
Earlier work this paper cites.
Actor-critic algorithms
Vijay Konda and John Tsitsiklis · 1999
Earlier work this paper cites.
Meta-semi: A meta-learning approach for semi-supervised learning
Yulin Wang, Jiayi Guo, Shiji Song, and Gao Huang · 2007
Earlier work this paper cites.
Reverse-mode ad in a functional framework: Lambda the ultimate backpropagator
Barak A. Pearlmutter and Jeffrey Mark Siskind · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Gradient-based hyperparameter optimization through reversible learning
Dougal Maclaurin, David Duvenaud, and Ryan Adams · 2015
Earlier work this paper cites.
Learning from massive noisy labeled data for image classification
Tong Xiao, Tian Xia, Yi Yang, Chang Huang, and Xiaogang Wang · 2015
Earlier work this paper cites.
On the convergence of stochastic bi-level gradient methods
Nicolas Couellan and Wenjuan Wang · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Data programming: Creating large training sets, quickly
Alexander J Ratner, Christopher M De Sa, Sen Wu, Daniel Selsam, and Christopher Ré · 2016
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Forward and reverse gradient-based hyperparameter optimization
Luca Franceschi, Michele Donini, Paolo Frasconi, and Massimiliano Pontil · 2017
Cited alongside, same era.
Deep hashing network for unsupervised domain adaptation
Hemanth Venkateswara, Jose Eusebio, Shayok Chakraborty, and Sethuraman Panchanathan · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
A new hyperparameters optimization method for convolutional neural networks
Hua Cui and Jie Bai · 2019
Cited alongside, same era.
Generalized inner loop meta-learning
Edward Grefenstette, Brandon Amos, Denis Yarats, Phu Mon Htut, Artem Molchanov, Franziska Meier, Douwe Kiela, Kyunghyun Cho, and Soumith Chintala · 2019
Cited alongside, same era.
A game theoretic framework for model based reinforcement learning
Aravind Rajeswaran, Igor Mordatch, and Vikash Kumar · 2020
Later among the works it cites.
Not all unlabeled data are equal: Learning to weight data in semi-supervised learning
Zhongzheng Ren, Raymond Yeh, and Alexander Schwing · 2020
Later among the works it cites.
Generative teaching networks: Accelerating neural architecture search by learning to generate synthetic training data
Felipe Petroski Such, Aditya Rawal, Joel Lehman, Kenneth Stanley, and Jeffrey Clune · 2020
Later among the works it cites.
Long-tailed classification by keeping the good and removing the bad momentum causal effect
Kaihua Tang, Jianqiang Huang, and Hanwang Zhang · 2020
Later among the works it cites.
How important is the train-validation split in meta-learning?
Yu Bai, Minshuo Chen, Pan Zhou, Tuo Zhao, Jason Lee, Sham Kakade, Huan Wang, and Caiming Xiong · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
DARTS: differentiable architecture search
Hanxiao Liu, Karen Simonyan, and Yiming Yang · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al · 2019
Cited alongside, same era.
Meta-learning with implicit gradients
Aravind Rajeswaran, Chelsea Finn, Sham M Kakade, and Sergey Levine · 2019
Cited alongside, same era.
Meta-weight-net: Learning an explicit mapping for sample weighting
Jun Shu, Qi Xie, Lixuan Yi, Qian Zhao, Sanping Zhou, Zongben Xu, and Deyu Meng · 2019
Cited alongside, same era.
learn2learn: A library for meta-learning research
Sébastien MR Arnold, Praateek Mahajan, Debajyoti Datta, Ian Bunner, and Konstantinos Saitas Zarkias · 2020
Cited alongside, same era.
On the iteration complexity of hypergradient computation
Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil, and Saverio Salzo · 2020
Cited alongside, same era.
Momentum contrast for unsupervised visual representation learning
Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross Girshick · 2020
Cited alongside, same era.
Mathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig, Stephan Hoyer, Felipe Llinares-López, Fabian Pedregosa, and Jean-Philippe Vert · 2021
Later among the works it cites.
Towards visual question answering on pathology images
Xuehai He, Zhuo Cai, Wenlan Wei, Yichen Zhang, Luntian Mou, Eric Xing, and Pengtao Xie · 2021
Later among the works it cites.
Bilevel optimization: Convergence analysis and enhanced design
Kaiyi Ji, Junjie Yang, and Yingbin Liang · 2021
Later among the works it cites.
Towards gradient-based bilevel optimization with non-convex followers and beyond
Risheng Liu, Yaohua Liu, Shangzhi Zeng, and Jin Zhang · 2021
Later among the works it cites.
Meta-learning to improve pre-training
Aniruddh Raghu, Jonathan Lorraine, Simon Kornblith, Matthew McDermott, and David K Duvenaud · 2021
Later among the works it cites.
A gradient method for multilevel optimization
Ryo Sato, Mirai Tanaka, and Akiko Takeda · 2021
Later among the works it cites.
Learning from mistakes–a framework for neural architecture search
Bhanu Garg, Li Zhang, Pradyumna Sridhara, Ramtin Hosseini, Eric Xing, and Pengtao Xie · 2022
Closest in time.
A multi-level optimization framework for end-to-end text augmentation
Sai Ashish Somayajula, Linfeng Song, and Pengtao Xie · 2022
Closest in time.
Performance-aware mutual knowledge distillation for improving neural architecture search
Pengtao Xie and Xuefeng Du · 2022
Closest in time.