Fetching the paper…
Reading the bibliography…
Generative Adversarial Networks are notoriously challenging to train.
Theory of games and economic behavior
J Von Neumann and O Morgenstern · 1944
Earlier work this paper cites.
The extragradient method for finding saddle points and other problems
Galina Michailovna Korpelevich · 1976
Earlier work this paper cites.
On the weak convergence of an ergodic iteration for the solution of variational inequalities for monotone operators in hilbert space
Ronald E Bruck · 1977
Earlier work this paper cites.
A modification of the arrow–hurwicz method for search of saddle points
Popov Leonid Denisovich · 1980
Earlier work this paper cites.
Efficient estimations from a slowly convergent Robbins-Monro process
David Ruppert · 1988
Earlier work this paper cites.
Nonlinear Differential Equations and Dynamical Systems
Ferdinand Verhulst · 1990
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
B. T. Polyak and A. B. Juditsky · 1992
Earlier work this paper cites.
Potential games
Dov Monderer and Lloyd S Shapley · 1996
Earlier work this paper cites.
The MNIST database of handwritten digits
Yann Lecun and Corinna Cortes · 1998
Earlier work this paper cites.
Nonlinear programming
Dimitri P Bertsekas · 1999
Earlier work this paper cites.
Finite-Dimensional Variational Inequalities and Complementarity Problems Vol I
Francisco Facchinei and Jong-Shi Pang · 2003
Earlier work this paper cites.
Learning Multiple Layers of Features from Tiny Images
Alex Krizhevsky · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y. Ng · 2011
Earlier work this paper cites.
Online learning with predictable sequences
Alexander Rakhlin and Karthik Sridharan · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jonathon Shlens, and Zbigniew Wojna · 2015
Cited alongside, same era.
NIPS 2016 tutorial: Generative adversarial networks
Ian Goodfellow · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, and Soumith Chintala · 2016
Cited alongside, same era.
Improved techniques for training GANs
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Cited alongside, same era.
Stabilizing adversarial nets with prediction methods
Abhay Yadav, Sohil Shah, Zheng Xu, David Jacobs, and Tom Goldstein · 2018
Later among the works it cites.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2019
Later among the works it cites.
Reducing noise in GAN training with variance reduced extragradient
Tatjana Chavdarova, Gauthier Gidel, François Fleuret, and Simon Lacoste-Julien · 2019
Later among the works it cites.
On the ineffectiveness of variance reduced optimization for deep learning
Aaron Defazio and Léon Bottou · 2019
Later among the works it cites.
Gradient Noise Convolution (GNC): Smoothing Loss Function for Distributed Large-Batch SGD
Kosuke Haruki, Taiji Suzuki, Yohei Hamakawa, Takeshi Toda, Ryuji Sakai, Masahiro Ozawa, and Mitsuhiro Kimura · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aymeric Dieuleveut, Alain Durmus, and Francis Bach · 2017
Cited alongside, same era.
GANs trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Cited alongside, same era.
The numerics of GANs
Lars Mescheder, Sebastian Nowozin, and Andreas Geiger · 2017
Cited alongside, same era.
Unrolled generative adversarial networks
Luke Metz, Ben Poole, David Pfau, and Jascha Sohl-Dickstein · 2017
Cited alongside, same era.
Deep decentralized multi-task multi-agent reinforcement learning under partial observability
Shayegan Omidshafiei, Jason Pazis, Chris Amato, Jonathan P. How, and John Vian · 2017
Cited alongside, same era.
The mechanics of n-player differentiable games
David Balduzzi, Sebastien Racaniere, James Martens, Jakob Foerster, Karl Tuyls, and Thore Graepel · 2018
Cited alongside, same era.
Training GANs with optimism
Constantinos Daskalakis, Andrew Ilyas, Vasilis Syrgkanis, and Haoyang Zeng · 2018
Cited alongside, same era.
Aryan Mokhtari, Asuman Ozdaglar, and Sarath Pattathil · 2019
Later among the works it cites.
Generative modeling by estimating gradients of the data distribution
Yang Song and Stefano Ermon · 2019
Later among the works it cites.
Local SGD converges fast and communicates little
Sebastian U. Stich · 2019
Later among the works it cites.
The unusual effectiveness of averaging in GAN training
Yasin Yazıcı, Chuan-Sheng Foo, Stefan Winkler, Kim-Hui Yap, Georgios Piliouras, and Vijay Chandrasekhar · 2019
Later among the works it cites.
Lookahead optimizer: k steps forward, 1 step back
Michael Zhang, James Lucas, Jimmy Ba, and Geoffrey E Hinton · 2019
Later among the works it cites.
A closer look at the optimization landscapes of generative adversarial networks
Hugo Berard, Gauthier Gidel, Amjad Almahairi, Pascal Vincent, and Simon Lacoste-Julien · 2020
Closest in time.
The complexity of constrained min-max optimization
Constantinos Daskalakis, Stratis Skoulakis, and Manolis Zampetakis · 2020
Closest in time.
What is local optimality in nonconvex-nonconcave minimax optimization?
Chi Jin, Praneeth Netrapalli, and Michael I. Jordan · 2020
Closest in time.
A unified theory of decentralized SGD with changing topology and local updates
Anastasia Koloskova, Nicolas Loizou, Sadra Boreiri, Martin Jaggi, and Sebastian U. Stich · 2020
Closest in time.
On the variance of the adaptive learning rate and beyond
Liyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen, Xiaodong Liu, Jianfeng Gao, and Jiawei Han · 2020
Closest in time.
Implicit learning dynamics in stackelberg games: Equilibria characterization, convergence analysis, and empirical study
Fiez Tanner, Chasnov Benjamin, and Ratliff Lillian · 2020
Closest in time.
Lookahead converges to stationary points of smooth non-convex functions
J. Wang, V. Tantia, N. Ballas, and M. Rabbat · 2020
Closest in time.
On solving minimax optimization locally: A follow-the-ridge approach
Yuanhao Wang, Guodong Zhang, and Jimmy Ba · 2020
Closest in time.
Understanding and stabilizing gans’ training dynamics with control theory
Kun Xu, Chongxuan Li, Huanshu Wei, Jun Zhu, and Bo Zhang · 2020
Closest in time.