Fetching the paper…
Reading the bibliography…
We develop an approach for text-to-image generation that embraces additional retrieval images, driven by a combination of implicit visual guidance loss and generative objectives.
A Simple Framework for Contrastive Learning of Visual Representations
Chen, T.; Kornblith, S.; Norouzi, M.; and Hinton, G. E. 2020a · 2002
Earlier work this paper cites.
Links between perceptrons, MLPs and SVMs
Collobert, R.; and Bengio, S. 2004 · 2004
Earlier work this paper cites.
DF-GAN: Deep Fusion Generative Adversarial Networks for Text-to-Image Synthesis
Tao, M.; Tang, H.; Wu, S.; Sebe, N.; Wu, F.; and Jing, X. 2020 · 2008
Earlier work this paper cites.
Image Generators with Conditionally-Independent Pixel Synthesis
Anokhin, I.; Demochkin, K.; Khakhulin, T.; Sterkin, G.; Lempitsky, V.; and Korzhenkov, D. 2020 · 2011
Earlier work this paper cites.
Adversarial Generation of Continuous Images
Skorokhodov, I.; Ignatyev, S.; and Elhoseiny, M. 2020 · 2011
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
Wah, C.; Branson, S.; Welinder, P.; Perona, P.; and Belongie, S. 2011 · 2011
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I. J.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A. C.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Unifying Visual-Semantic Embeddings with Multimodal Neural Language Models
Kiros, R.; Salakhutdinov, R.; and Zemel, R. S. 2014 · 2014
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions
Karpathy, A.; and Li, F. 2015 · 2015
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
You Only Look Once: Unified, Real-Time Object Detection
Redmon, J.; Divvala, S. K.; Girshick, R. B.; and Farhadi, A. 2016 · 2016
Earlier work this paper cites.
Improved Techniques for Training GANs
Salimans, T.; Goodfellow, I. J.; Zaremba, W.; Cheung, V.; Radford, A.; and Chen, X. 2016 · 2016
Earlier work this paper cites.
Learning Deep Structure-Preserving Image-Text Embeddings
Wang, L.; Li, Y.; and Lazebnik, S. 2016 · 2016
Earlier work this paper cites.
An Empirical Study of Language CNN for Image Captioning
Gu, J.; Wang, G.; Cai, J.; and Chen, T. 2017 · 2017
Earlier work this paper cites.
HyperNetworks
Ha, D.; Dai, A. M.; and Le, Q. V. 2017 · 2017
Cited alongside, same era.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Cited alongside, same era.
StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks
Zhang, H.; Xu, T.; Li, H.; Zhang, S.; Wang, X.; Huang, X.; and Metaxas, D. 2017 · 2017
Cited alongside, same era.
Look, Imagine and Match: Improving Textual-Visual Cross-Modal Retrieval With Generative Models
Gu, J.; Cai, J.; Joty, S. R.; Niu, L.; and Wang, G. 2018 · 2018
Cited alongside, same era.
Conceptual Captions: A Cleaned, Hypernymed, Image Alt-text Dataset For Automatic Image Captioning
Sharma, P.; Ding, N.; Goodman, S.; and Soricut, R. 2018 · 2018
Cited alongside, same era.
Self-Supervised Relationship Probing
Gu, J.; Kuen, J.; Joty, S.; Cai, J.; Morariu, V.; Zhao, H.; and Sun, T. 2020 · 2020
Later among the works it cites.
Semantic Object Accuracy for Generative Text-to-Image Synthesis
Hinz, T.; Heinrich, S.; and Wermter, S. 2020 · 2020
Later among the works it cites.
Analyzing and Improving the Image Quality of StyleGAN
Karras, T.; Laine, S.; Aittala, M.; Hellsten, J.; Lehtinen, J.; and Aila, T. 2020 · 2020
Later among the works it cites.
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2020 · 2020
Later among the works it cites.
Instance-Conditioned GAN
Casanova, A.; Careil, M.; Verbeek, J.; Drozdzal, M.; and Romero-Soriano, A. 2021 · 2021
Later among the works it cites.
Text-to-Image Generation Grounded by Fine-Grained User Attention
Koh, J. Y.; Baldridge, J.; Lee, H.; and Yang, Y. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xu, T.; Zhang, P.; Huang, Q.; Zhang, H.; Gan, Z.; Huang, X.; and He, X. 2018 · 2018
Cited alongside, same era.
Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?
Abdal, R.; Qin, Y.; and Wonka, P. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Generating Multiple Objects at Spatially Distinct Locations
Hinz, T.; Heinrich, S.; and Wermter, S. 2019 · 2019
Cited alongside, same era.
A Style-Based Generator Architecture for Generative Adversarial Networks
Karras, T.; Laine, S.; and Aila, T. 2019 · 2019
Cited alongside, same era.
ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
Lu, J.; Batra, D.; Parikh, D.; and Lee, S. 2019 · 2019
Cited alongside, same era.
stylegan-encoder
Nikitko, D. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery
Patashnik, O.; Wu, Z.; Shechtman, E.; Cohen-Or, D.; and Lischinski, D. 2021 · 2021
Later among the works it cites.
Learning Transferable Visual Models From Natural Language Supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I. 2021 · 2021
Later among the works it cites.
Zero-Shot Text-to-Image Generation
Ramesh, A.; Pavlov, M.; Goh, G.; Gray, S.; Voss, C.; Radford, A.; Chen, M.; and Sutskever, I. 2021 · 2021
Later among the works it cites.
Cross-Modal Contrastive Learning for Text-to-Image Generation
Zhang, H.; Koh, J. Y.; Baldridge, J.; Lee, H.; and Yang, Y. 2021 · 2021
Later among the works it cites.
Stylizing 3D Scene via Implicit Representation and HyperNetwork
Chiang, P.-Z.; Tsai, M.-S.; Tseng, H.-Y.; Lai, W.-S.; and Chiu, W.-C. 2022 · 2022
Closest in time.
HyperCGAN: Text-to-Image Synthesis with HyperNet-Modulated Conditional Generative Adversarial Networks
HAYDAROV, K.; Muhamed, A.; Lazarevic, J.; Skorokhodov, I.; and Elhoseiny, M. 2022 · 2022
Closest in time.
Memory-Driven Text-to-Image Generation
Li, B.; Torr, P.; and Lukasiewicz, T. 2022 · 2022
Closest in time.