Fetching the paper…
Reading the bibliography…
We propose Styleformer, which is a style-based generator for GAN architecture, but a convolution-free transformer-based generator.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
An analysis of single-layer networks in unsupervised feature learning
Adam Coates, Andrew Ng, and Honglak Lee · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Generative adversarial networks, 2014
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding, 2016
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning, 2016
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C. Lawrence Zitnick, and Ross Girshick · 2016
Earlier work this paper cites.
Unsupervised representation learning with deep convolutional generative adversarial networks, 2016
Alec Radford, Luke Metz, and Soumith Chintala · 2016
Earlier work this paper cites.
Improved techniques for training gans, 2016
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Earlier work this paper cites.
Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop, 2016
Fisher Yu, Ari Seff, Yinda Zhang, Shuran Song, Thomas Funkhouser, and Jianxiong Xiao · 2016
Earlier work this paper cites.
Wasserstein gan, 2017
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Earlier work this paper cites.
A learned representation for artistic style, 2017
Vincent Dumoulin, Jonathon Shlens, and Manjunath Kudlur · 2017
Earlier work this paper cites.
Exploring the structure of a real-time, arbitrary neural artistic stylization network, 2017
Golnaz Ghiasi, Honglak Lee, Manjunath Kudlur, Vincent Dumoulin, and Jonathon Shlens · 2017
Earlier work this paper cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications, 2017
Andrew G. Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam · 2017
Earlier work this paper cites.
Arbitrary style transfer in real-time with adaptive instance normalization, 2017
Xun Huang and Serge Belongie · 2017
Earlier work this paper cites.
Photo-realistic single image super-resolution using a generative adversarial network, 2017
Christian Ledig, Lucas Theis, Ferenc Huszar, Jose Caballero, Andrew Cunningham, Alejandro Acosta, Andrew Aitken, Alykhan Tejani, Johannes Totz, Zehan Wang, and Wenzhe Shi · 2017
Earlier work this paper cites.
Least squares generative adversarial networks, 2017
Xudong Mao, Qing Li, Haoran Xie, Raymond Y. K. Lau, Zhen Wang, and Stephen Paul Smolley · 2017
Earlier work this paper cites.
Attention is all you need, 2017
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Escaping from collapsing modes in a constrained space, 2018
Chia-Che Chang, Chieh Hubert Lin, Che-Rung Lee, Da-Cheng Juan, Wei Wei, and Hwann-Tzong Chen · 2018
Earlier work this paper cites.
Stargan: Unified generative adversarial networks for multi-domain image-to-image translation, 2018
Yunjey Choi, Minje Choi, Munyoung Kim, Jung-Woo Ha, Sunghun Kim, and Jaegul Choo · 2018
Cited alongside, same era.
Feature-wise transformations
Vincent Dumoulin, Ethan Perez, Nathan Schucher, Florian Strub, Harm de Vries, Aaron Courville, and Yoshua Bengio · 2018
Cited alongside, same era.
Gans trained by a two time-scale update rule converge to a local nash equilibrium, 2018
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2018
Cited alongside, same era.
Image-to-image translation with conditional adversarial networks, 2018
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros · 2018
Cited alongside, same era.
Image generation from scene graphs, 2018
Justin Johnson, Agrim Gupta, and Li Fei-Fei · 2018
Cited alongside, same era.
Progressive growing of gans for improved quality, stability, and variation, 2018
Analyzing and improving the image quality of stylegan, 2020
Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila · 2020
Later among the works it cites.
Big transfer (bit): General visual representation learning, 2020
Alexander Kolesnikov, Lucas Beyer, Xiaohua Zhai, Joan Puigcerver, Jessica Yung, Sylvain Gelly, and Neil Houlsby · 2020
Later among the works it cites.
Top-k training of gans: Improving gan performance by throwing away bad samples, 2020
Samarth Sinha, Zhengli Zhao, Anirudh Goyal, Colin Raffel, and Augustus Odena · 2020
Later among the works it cites.
Discriminator contrastive divergence: Semi-amortized generative modeling by exploring energy of the discriminator, 2020
Yuxuan Song, Qiwei Ye, Minkai Xu, and Tie-Yan Liu · 2020
Later among the works it cites.
Efficientnet: Rethinking model scaling for convolutional neural networks, 2020
Mingxing Tan and Quoc V. Le · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen · 2018
Cited alongside, same era.
Spectral normalization for generative adversarial networks, 2018
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida · 2018
Cited alongside, same era.
Generative image inpainting with contextual attention, 2018
Jiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen, Xin Lu, and Thomas S. Huang · 2018
Cited alongside, same era.
Style generator inversion for image enhancement and animation, 2019
Aviv Gabbay and Yedid Hoshen · 2019
Cited alongside, same era.
Autogan: Neural architecture search for generative adversarial networks, 2019
Xinyu Gong, Shiyu Chang, Yifan Jiang, and Zhangyang Wang · 2019
Cited alongside, same era.
A style-based generator architecture for generative adversarial networks, 2019
Tero Karras, Samuli Laine, and Timo Aila · 2019
Cited alongside, same era.
Improving mmd-gan training with repulsive loss function, 2019
Wei Wang, Yuan Sun, and Saman Halgamuge · 2019
Cited alongside, same era.
Linformer: Self-attention with linear complexity, 2020
Sinong Wang, Belinda Z. Li, Madian Khabsa, Han Fang, and Hao Ma · 2020
Later among the works it cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks, 2020
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros · 2020
Later among the works it cites.
Sean: Image synthesis with semantic region-adaptive normalization
Peihao Zhu, Rameen Abdal, Yipeng Qin, and Peter Wonka · 2020
Later among the works it cites.
Semantically multi-modal image synthesis, 2020
Zhen Zhu, Zhiliang Xu, Ansheng You, and Xiang Bai · 2020
Later among the works it cites.
Mobilestylegan: A lightweight convolutional neural network for high-fidelity image synthesis, 2021
Sergei Belousov · 2021
Closest in time.
Is space-time attention all you need for video understanding?, 2021
Gedas Bertasius, Heng Wang, and Lorenzo Torresani · 2021
Closest in time.
Taming transformers for high-resolution image synthesis, 2021
Patrick Esser, Robin Rombach, and Björn Ommer · 2021
Closest in time.
Levit: a vision transformer in convnet’s clothing for faster inference, 2021
Ben Graham, Alaaeldin El-Nouby, Hugo Touvron, Pierre Stock, Armand Joulin, Hervé Jégou, and Matthijs Douze · 2021
Closest in time.
Generative adversarial transformers, 2021
Drew A. Hudson and C. Lawrence Zitnick · 2021
Closest in time.
Transgan: Two transformers can make one strong gan, 2021
Yifan Jiang, Shiyu Chang, and Zhangyang Wang · 2021
Closest in time.
Swin transformer: Hierarchical vision transformer using shifted windows, 2021
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Closest in time.
Mlp-mixer: An all-mlp architecture for vision, 2021
Ilya Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer, Xiaohua Zhai, Thomas Unterthiner, Jessica Yung, Andreas Steiner, Daniel Keysers, Jakob Uszkoreit, Mario Lucic, and Alexey Dosovitskiy · 2021
Closest in time.
Peergan: Generative adversarial networks with a competing peer discriminator, 2021
Jiaheng Wei, Minghao Liu, Jiahao Luo, Qiutong Li, James Davis, and Yang Liu · 2021
Closest in time.
Cvt: Introducing convolutions to vision transformers, 2021
Haiping Wu, Bin Xiao, Noel Codella, Mengchen Liu, Xiyang Dai, Lu Yuan, and Lei Zhang · 2021
Closest in time.
Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers, 2021
Sixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu, Zekun Luo, Yabiao Wang, Yanwei Fu, Jianfeng Feng, Tao Xiang, Philip H. S. Torr, and Li Zhang · 2021
Closest in time.