Fetching the paper…
Reading the bibliography…
Although two-stage Vector Quantized (VQ) generative models allow for synthesizing high-fidelity and high-resolution images, their quantization operator encodes similar patches within an image into the same index, resulting in a repeated artifact for similar adjacent regions using existing decoder architectures.
Nice: Non-linear independent components estimation
L. Dinh, D. Krueger, and Y. Bengio · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
J. Sohl-Dickstein, E. A. Weiss, N. Maheswaranathan, and S. Ganguli · 2015
Earlier work this paper cites.
A learned representation for artistic style
V. Dumoulin, J. Shlens, and M. Kudlur · 2016
Earlier work this paper cites.
Improved techniques for training gans
T. Salimans, I. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen · 2016
Earlier work this paper cites.
Conditional image generation with pixelcnn decoders
A. Van den Oord, N. Kalchbrenner, L. Espeholt, O. Vinyals, A. Graves, et al · 2016
Earlier work this paper cites.
Convolutional sequence to sequence learning
J. Gehring, M. Auli, D. Grangier, D. Yarats, and Y. N. Dauphin · 2017
Earlier work this paper cites.
Improved training of wasserstein gans
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. C. Courville · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter · 2017
Earlier work this paper cites.
Arbitrary style transfer in real-time with adaptive instance normalization
X. Huang and S. Belongie · 2017
Earlier work this paper cites.
Neural discrete representation learning
A. Van Den Oord, O. Vinyals, et al · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Large scale gan training for high fidelity natural image synthesis
A. Brock, J. Donahue, and K. Simonyan · 2018
Earlier work this paper cites.
Multimodal unsupervised image-to-image translation
X. Huang, M.-Y. Liu, S. Belongie, and J. Kautz · 2018
Earlier work this paper cites.
Glow: Generative flow with invertible 1x1 convolutions
D. P. Kingma and P. Dhariwal · 2018
Cited alongside, same era.
Group normalization
Y. Wu and K. He · 2018
Cited alongside, same era.
The unreasonable effectiveness of deep features as a perceptual metric
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Cited alongside, same era.
A style-based generator architecture for generative adversarial networks
T. Karras, S. Laine, and T. Aila · 2019
Cited alongside, same era.
Semantic image synthesis with spatially-adaptive normalization
T. Park, M.-Y. Liu, T.-C. Wang, and J.-Y. Zhu · 2019
Cited alongside, same era.
Alias-free generative adversarial networks
T. Karras, M. Aittala, S. Laine, E. Härkönen, J. Hellsten, J. Lehtinen, and T. Aila · 2021
Later among the works it cites.
Generating images with sparse representations
C. Nash, J. Menick, S. Dieleman, and P. Battaglia · 2021
Later among the works it cites.
Improved denoising diffusion probabilistic models
A. Q. Nichol and P. Dhariwal · 2021
Later among the works it cites.
Latent video transformer
R. Rakhimov, D. Volkhonskiy, A. Artemov, D. Zorin, and E. Burnaev · 2021
Later among the works it cites.
Zero-shot text-to-image generation
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, M. Chen, and I. Sutskever · 2021
Later among the works it cites.
High-resolution image synthesis with latent diffusion models, 2021
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Generating diverse high-fidelity images with vq-vae-2
A. Razavi, A. Van den Oord, and O. Vinyals · 2019
Cited alongside, same era.
Generative pretraining from pixels
M. Chen, A. Radford, R. Child, J. Wu, H. Jun, D. Luan, and I. Sutskever · 2020
Cited alongside, same era.
Analyzing and improving the image quality of stylegan
T. Karras, S. Laine, M. Aittala, J. Hellsten, J. Lehtinen, and T. Aila · 2020
Cited alongside, same era.
Fourier features let networks learn high frequency functions in low dimensional domains
M. Tancik, P. Srinivasan, B. Mildenhall, S. Fridovich-Keil, N. Raghavan, U. Singhal, R. Ramamoorthi, J. Barron, and R. Ng · 2020
Cited alongside, same era.
NVAE: A deep hierarchical variational autoencoder
A. Vahdat and J. Kautz · 2020
Cited alongside, same era.
Navigating the gan parameter space for semantic image editing
A. Cherepkov, A. Voynov, and A. Babenko · 2021
Cited alongside, same era.
C. Wu, J. Liang, L. Ji, F. Yang, Y. Fang, D. Jiang, and N. Duan · 2021
Later among the works it cites.
Videogpt: Video generation using vq-vae and transformers
W. Yan, Y. Zhang, P. Abbeel, and A. Srinivas · 2021
Later among the works it cites.
Maskgit: Masked generative image transformer
H. Chang, H. Zhang, L. Jiang, C. Liu, and W. T. Freeman · 2022
Closest in time.
Global context with discrete diffusion in vector quantised modelling for image generation
M. Hu, Y. Wang, T.-J. Cham, J. Yang, and P. Suganthan · 2022
Closest in time.
Autoregressive image generation using residual quantization
D. Lee, C. Kim, S. Kim, M. Cho, and W.-S. Han · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
A. Ramesh, P. Dhariwal, A. Nichol, C. Chu, and M. Chen · 2022
Closest in time.
Improving gan equilibrium by raising spatial awareness
J. Wang, C. Yang, Y. Xu, Y. Shen, H. Li, and B. Zhou · 2022
Closest in time.
Vector-quantized image modeling with improved VQGAN
J. Yu, X. Li, J. Y. Koh, H. Zhang, R. Pang, J. Qin, A. Ku, Y. Xu, J. Baldridge, and Y. Wu · 2022
Closest in time.
High-quality pluralistic image completion via code shared vqgan
C. Zheng, G. Song, T.-J. Cham, J. Cai, D. Phung, and L. Luo · 2022
Closest in time.