Fetching the paper…
Reading the bibliography…
With the advent of generative adversarial networks, synthesizing images from textual descriptions has recently become an active research area.
Z. S. Harris, Distributional structure, Word 10 (2–3) (1954) 146–162
1954
Earlier work this paper cites.
H. Zhang, T. Xu, H. Li, S. Zhang, X. Wang, X. Huang, D. N. Metaxas, Stackgan++: Realistic image synthesis with stacked generative adversarial networks, IEEE Transactions on Pattern Analysis and Machine Intelligence 41 (2017) 1947–1962
1962
Earlier work this paper cites.
J. Bromley, J. W. Bentz, L. Bottou, I. Guyon, Y. LeCun, C. Moore, E. Säckinger, R. Shah, Signature verification using a “siamese” time delay neural network, International Journal of Pattern Recognition and Artificial Intelligence 7 (04) (1993) 669–688
1993
Earlier work this paper cites.
M. Schuster, K. K. Paliwal, Bidirectional recurrent neural networks, IEEE Transactions on Signal Processing 45 (11) (1997) 2673–2681
1997
Earlier work this paper cites.
S. Kosslyn, G. Ganis, W. L. Thompson, Neural foundations of imagery, Nature Reviews Neuroscience 2 (2001) 635–642
2001
Earlier work this paper cites.
Y. Balaji, M. R. Min, B. Bai, R. Chellappa, H. P. Graf, Conditional gan with discriminative filter generation for text-to-video synthesis, in: Proceedings of the International Joint Conference on Artificial Intelligence, 2019, pp. 1995–2001
2001
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, W.-J. Zhu, Bleu: a method for automatic evaluation of machine translation, in: Proceedings of the 40th annual meeting of the Association for Computational Linguistics, 2002, pp. 311–318
2002
Earlier work this paper cites.
2002
Earlier work this paper cites.
S. Chopra, R. Hadsell, Y. LeCun, Learning a similarity metric discriminatively, with application to face verification, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2005, pp. 539–546
2005
Earlier work this paper cites.
A. Hyvärinen, Estimation of non-normalized statistical models by score matching, Journal of Machine Learning Research 6 (Apr) (2005) 695–709
2005
Earlier work this paper cites.
R. Hadsell, S. Chopra, Y. LeCun, Dimensionality reduction by learning an invariant mapping, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2006, pp. 1735–1742
2006
Earlier work this paper cites.
A. Lavie, A. Agarwal, Meteor: An automatic metric for mt evaluation with high levels of correlation with human judgments, in: Proceedings of the Second Workshop on Statistical Machine Translation, 2007, pp. 228–231
2007
Earlier work this paper cites.
M.-E. Nilsback, A. Zisserman, Automated flower classification over a large number of classes, in: Indian Conference on Computer Vision, Graphics & Image Processing, 2008, pp. 722–729
2008
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, J. Weston, Curriculum learning, in: International Conference on Machine Learning, 2009, pp. 41–48
2009
Earlier work this paper cites.
Y. LeCun, C. Cortes, C. Burges, Mnist handwritten digit database (2010)
2010
Earlier work this paper cites.
C. Wah, S. Branson, P. Welinder, P. Perona, S. Belongie, The caltech-ucsd birds-200-2011 dataset, California Institute of Technology (2011)
2011
Earlier work this paper cites.
M. Eitz, J. Hays, M. Alexa, How do humans sketch objects?, ACM Transactions on Graphics 31 (4) (2012) 1–10
2012
Earlier work this paper cites.
Y. Bengio, A. C. Courville, P. Vincent, Representation learning: A review and new perspectives, IEEE Transactions on Pattern Analysis and Machine Intelligence 35 (2013) 1798–1828
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, J. Dean, Distributed representations of words and phrases and their compositionality, in: Advances in Neural Information Processing Systems, 2013, pp. 3111–3119
2013
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. C. Courville, Y. Bengio, Generative adversarial nets, in: Advances in Neural Information Processing Systems, 2014, pp. 2672–2680
2014
Earlier work this paper cites.
M. Mirza, S. Osindero, Conditional generative adversarial nets, arXiv:1411.1784 (2014)
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. J. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, C. L. Zitnick, Microsoft coco: Common objects in context, in: European Conference on Computer Vision, 2014, pp. 740–755
2014
Earlier work this paper cites.
J. Pennington, R. Socher, C. D. Manning, Glove: Global vectors for word representation, in: Proceedings of the Conference on Empirical Methods in Natural Language Processing, 2014, pp. 1532––1543
2014
Earlier work this paper cites.
D. P. Kingma, M. Welling, Auto-encoding variational bayes, in: International Conference on Learning Representations, 2014
2014
Earlier work this paper cites.
R. Kiros, Y. Zhu, R. Salakhutdinov, R. S. Zemel, R. Urtasun, A. Torralba, S. Fidler, Skip-thought vectors, in: Advances in Neural Information Processing Systems, 2015, pp. 3294–3302
2015
Earlier work this paper cites.
K. Simonyan, A. Zisserman, Very deep convolutional networks for large-scale image recognition, in: International Conference on Learning Representations, 2015
2015
Earlier work this paper cites.
D. Bahdanau, K. Cho, Y. Bengio, Neural machine translation by jointly learning to align and translate, in: International Conference on Learning Representations, 2015
2015
Earlier work this paper cites.
T. Luong, H. Pham, C. D. Manning, Effective approaches to attention-based neural machine translation, in: Proceedings of the Conference on Empirical Methods in Natural Language Processing, 2015, pp. 1412–1421
2015
Earlier work this paper cites.
A. Karpathy, F.-F. Li, Deep visual-semantic alignments for generating image descriptions, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2015, pp. 3128–3137
2015
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, D. Erhan, Show and tell: A neural image caption generator, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2015, pp. 3156–3164
2015
Earlier work this paper cites.
K. S. Tai, R. Socher, C. D. Manning, Improved semantic representations from tree-structured long short-term memory networks, in: Proceedings of the ACL and International Joint Conference on Natural Language Processing, 2015, pp. 1556–1566
2015
Earlier work this paper cites.
S. Sukhbaatar, A. Szlam, J. Weston, R. Fergus, End-to-end memory networks, in: Advances in Neural Information Processing Systems, 2015, pp. 2440––2448
2015
Earlier work this paper cites.
L. Dinh, D. Krueger, Y. Bengio, Nice: Non-linear independent components estimation, in: International Conference on Learning Representations (Workshop), 2015
2015
Earlier work this paper cites.
R. B. Girshick, Fast r-cnn, in: Proceedings of the IEEE International Conference on Computer Vision, 2015, pp. 1440–1448
2015
Earlier work this paper cites.
J. Johnson, R. Krishna, M. A. Stark, L.-J. Li, D. A. Shamma, M. S. Bernstein, F.-F. Li, Image retrieval using scene graphs, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2015, pp. 3668–3678
2015
Earlier work this paper cites.
X. Shi, Z. Chen, H. Wang, D.-Y. Yeung, W.-K. Wong, W. chun Woo, Convolutional lstm network: A machine learning approach for precipitation nowcasting, in: Advances in Neural Information Processing Systems, 2015, p. 802–810
2015
Earlier work this paper cites.
R. Vedantam, C. Lawrence Zitnick, D. Parikh, Cider: Consensus-based image description evaluation, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2015, pp. 4566–4575
2015
Earlier work this paper cites.
C. Ledig, L. Theis, F. Huszár, J. A. Caballero, A. Aitken, A. Tejani, J. Totz, Z. Wang, W. Shi, Photo-realistic single image super-resolution using a generative adversarial network, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 4681–4690
2016
Earlier work this paper cites.
R. A. Yeh, C. Chen, T.-Y. Lim, A. G. Schwing, M. Hasegawa-Johnson, M. N. Do, Semantic image inpainting with deep generative models, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 5485–5493
2016
Earlier work this paper cites.
L. A. Gatys, A. S. Ecker, M. Bethge, Image style transfer using convolutional neural networks, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 2414–2423
2016
Earlier work this paper cites.
P. Isola, J.-Y. Zhu, T. Zhou, A. A. Efros, Image-to-image translation with conditional adversarial networks, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 1125–1134
2016
Earlier work this paper cites.
S. E. Reed, Z. Akata, X. Yan, L. Logeswaran, B. Schiele, H. Lee, Generative adversarial text to image synthesis, in: International Conference on Machine Learning, 2016, pp. 1060–1069
2016
Earlier work this paper cites.
A. Odena, C. Olah, J. Shlens, Conditional image synthesis with auxiliary classifier gans, in: International Conference on Machine Learning, 2016, pp. 2642–2651
2016
Earlier work this paper cites.
S. Reed, Z. Akata, H. Lee, B. Schiele, Learning deep representations of fine-grained visual descriptions, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 49–58
2016
Earlier work this paper cites.
H. Zhang, T. Xu, H. Li, Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks, in: Proceedings of the IEEE International Conference on Computer Vision, 2016, pp. 5907–5915
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 770–778
2016
Earlier work this paper cites.
A. H. Miller, A. Fisch, J. Dodge, A.-H. Karimi, A. Bordes, J. Weston, Key-value memory networks for directly reading documents, in: Proceedings of the Conference on Empirical Methods in Natural Language Processing, 2016, pp. 1400––1409
2016
Earlier work this paper cites.
S. E. Reed, Z. Akata, S. Mohan, S. Tenka, B. Schiele, H. Lee, Learning what and where to draw, in: Advances in Neural Information Processing Systems, 2016, pp. 217–225
2016
Earlier work this paper cites.
S. E. Reed, A. van den Oord, N. Kalchbrenner, V. Bapst, M. M. Botvinick, N. de Freitas, Generating interpretable images with controllable structure, Technical Report (2016)
2016
Earlier work this paper cites.
A. van den Oord, N. Kalchbrenner, K. Kavukcuoglu, Pixel recurrent neural networks, in: International Conference on Machine Learning, 2016, pp. 1747–1756
2016
Earlier work this paper cites.
L. Theis, A. van den Oord, M. Bethge, A note on the evaluation of generative models, in: International Conference on Learning Representations, 2016
2016
Earlier work this paper cites.
T. Salimans, I. Goodfellow, W. Zaremba, V. Cheung, A. Radford, X. Chen, Improved techniques for training gans, in: Advances in Neural Information Processing Systems, 2016, pp. 2234–2242
2016
Earlier work this paper cites.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, Z. Wojna, Rethinking the inception architecture for computer vision, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2016, pp. 2818–2826
2016
Earlier work this paper cites.
A. van den Oord, N. Kalchbrenner, L. Espeholt, K. Kavukcuoglu, O. Vinyals, A. Graves, Conditional image generation with pixelcnn decoders, in: Advances in Neural Information Processing Systems, 2016, pp. 4790–4798
2016
Earlier work this paper cites.
J.-Y. Zhu, T. Park, P. Isola, A. A. Efros, Unpaired image-to-image translation using cycle-consistent adversarial networks, in: Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 2223–2232
2017
Earlier work this paper cites.
X. Wu, K. Xu, P. Hall, A survey of image synthesis and editing with generative adversarial networks, Tsinghua Science and Technology 22 (6) (2017) 660–674
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
T. Xu, P. Zhang, Q. Huang, H. Zhang, Z. Gan, X. Huang, X. He, Attngan: Fine-grained text to image generation with attentional generative adversarial networks, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2017, pp. 1316–1324
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. B. Girshick, K. He, B. Hariharan, S. J. Belongie, Feature pyramid networks for object detection, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2017, pp. 936–944
2017
Earlier work this paper cites.
W.-S. Lai, J.-B. Huang, N. Ahuja, M.-H. Yang, Deep laplacian pyramid networks for fast and accurate super-resolution, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2017, pp. 5835–5843
2017
Earlier work this paper cites.
C. Ledig, L. Theis, F. Huszár, J. Caballero, A. Cunningham, A. Acosta, A. Aitken, A. Tejani, J. Totz, Z. Wang, et al., Photo-realistic single image super-resolution using a generative adversarial network, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2017, pp. 4681–4690
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, I. Polosukhin, Attention is all you need, in: Advances in Neural Information Processing Systems, 2017, pp. 5998–6008
2017
Earlier work this paper cites.
Z. Lin, M. Feng, C. N. dos Santos, M. Yu, B. Xiang, B. Zhou, Y. Bengio, A structured self-attentive sentence embedding, in: International Conference on Learning Representations, 2017
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
A. Nguyen, J. Clune, Y. Bengio, A. Dosovitskiy, J. Yosinski, Plug & play generative networks: Conditional iterative generation of images in latent space, in: Advances in Neural Information Processing Systems, 2017, pp. 4467–4477
2017
Cited alongside, same era.
J. Donahue, P. Krähenbühl, T. Darrell, Adversarial feature learning, in: International Conference on Learning Representations, 2017
2017
Cited alongside, same era.
V. Dumoulin, I. Belghazi, B. Poole, A. Lamb, M. Arjovsky, O. Mastropietro, A. C. Courville, Adversarially learned inference, in: International Conference on Learning Representations, 2017
2017
Cited alongside, same era.
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, S. Lee, J. M. F. Moura, D. Parikh, D. Batra, Visual dialog, IEEE Transactions on Pattern Analysis and Machine Intelligence 41 (2019) 1242–1256
2019
Later among the works it cites.
T. Hinz, S. Heinrich, S. Wermter, Generating multiple objects at spatially distinct locations, in: International Conference on Learning Representations, 2019
2019
Later among the works it cites.
B. Zhao, L. Meng, W. Yin, L. Sigal, Image generation from layout, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2019, pp. 8584–8593
2019
Later among the works it cites.
W. Sun, T. Wu, Image synthesis from reconfigurable layout and style, in: Proceedings of the IEEE International Conference on Computer Vision, 2019, pp. 10531–10540
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X. Huang, S. Belongie, Arbitrary style transfer in real-time with adaptive instance normalization, in: Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 1501–1510
2017
Cited alongside, same era.
V. Dumoulin, I. Belghazi, B. Poole, A. Lamb, M. Arjovsky, O. Mastropietro, A. C. Courville, Adversarially learned inference, in: International Conference on Learning Representations, 2017
2017
Cited alongside, same era.
L. Dinh, J. Sohl-Dickstein, S. Bengio, Density estimation using real nvp, in: International Conference on Learning Representations, 2017
2017
Cited alongside, same era.
Y. Goyal, T. Khot, D. Summers-Stay, D. Batra, D. Parikh, Making the v in vqa matter: Elevating the role of image understanding in visual question answering, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2017, pp. 6325–6334
2017
Cited alongside, same era.
H. Ben-younes, R. Cadène, M. Cord, N. Thome, Mutan: Multimodal tucker fusion for visual question answering, in: Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 2631–2639
2017
Cited alongside, same era.
S. E. Reed, A. van den Oord, N. Kalchbrenner, S. G. Colmenarejo, Z. Wang, Y. Chen, D. Belov, N. de Freitas, Parallel multiscale autoregressive density estimation, in: International Conference on Machine Learning, 2017, pp. 2912–2921
2017
Cited alongside, same era.
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, et al., Visual genome: Connecting language and vision using crowdsourced dense image annotations, International Journal of Computer Vision 123 (1) (2017) 32–73
2017
Cited alongside, same era.
Q. Chen, V. Koltun, Photographic image synthesis with cascaded refinement networks, in: Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 1511–1520
2017
Cited alongside, same era.
2019
Later among the works it cites.
T. ting Qiao, J. Zhang, D. Xu, D. Tao, Learn, imagine and create: Text-to-image generation from prior knowledge, in: Advances in Neural Information Processing Systems, 2019, pp. 887—897
2019
Later among the works it cites.
O. Ashual, L. Wolf, Specifying object attributes and relations in interactive scene generation, in: Proceedings of the IEEE International Conference on Computer Vision, 2019, pp. 4561–4569
2019
Later among the works it cites.
Y. Li, T. Ma, Y. Bai, N. Duan, S. Wei, X. Wang, Pastegan: A semi-parametric method to generate image from scene graph, in: Advances in Neural Information Processing Systems, 2019
2019
Later among the works it cites.
G. Mittal, S. Agrawal, A. Agarwal, S. Mehta, T. Marwah, Interactive image generation using scene graphs, in: International Conference on Learning Representations (Workshop), 2019
2019
Later among the works it cites.
S. V. Ravuri, O. Vinyals, Classification accuracy score for conditional generative models, in: Advances in Neural Information Processing Systems, 2019, pp. 12268–12279
2019
Later among the works it cites.
J. Lu, D. Batra, D. Parikh, S. Lee, Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks, in: Advances in Neural Information Processing Systems, 2019, pp. 13–23
2019
Later among the works it cites.
2019
Later among the works it cites.
H. H. Tan, M. Bansal, Lxmert: Learning cross-modality encoder representations from transformers, in: Proceedings of the Conference on Empirical Methods in Natural Language Processing, 2019, pp. 5100––5111
2019
Later among the works it cites.
A. Odena, Open questions about generative adversarial networks, Distill (2019)
2019
Later among the works it cites.
A. Razavi, A. van den Oord, O. Vinyals, Generating diverse high-fidelity images with vq-vae-2, in: Advances in Neural Information Processing Systems, 2019, pp. 14866–14876
2019
Later among the works it cites.
2019
Later among the works it cites.
Y. Song, S. Ermon, Generative modeling by estimating gradients of the data distribution, in: Advances in Neural Information Processing Systems, 2019, pp. 11918–11930
2019
Later among the works it cites.
M. O. Turkoglu, L. Spreeuwers, W. Thong, B. Kicanaoglu, A layer-based sequential framework for scene generation with gans, in: Proceedings of the AAAI Conference on Artificial Intelligence, 2019, pp. 8901–8908
2019
Later among the works it cites.
A. El-Nouby, S. Sharma, H. Schulz, D. Hjelm, L. E. Asri, S. E. Kahou, Y. Bengio, G. W.Taylor, Tell, draw, and repeat: Generating and modifying images based on continual linguistic instruction, in: Proceedings of the IEEE International Conference on Computer Vision, 2019, pp. 10304–10312
2019
Later among the works it cites.
T. Kynkäänniemi, T. Karras, S. Laine, J. Lehtinen, T. Aila, Improved precision and recall metric for assessing generative models, in: Advances in Neural Information Processing Systems, 2019, pp. 3927–3936
2019
Later among the works it cites.
2019
Later among the works it cites.
S. Zhou, M. Gordon, R. Krishna, A. Narcomey, L. F. Fei-Fei, M. Bernstein, Hype: A benchmark for human eye perceptual evaluation of generative models, in: Advances in Neural Information Processing Systems, 2019, pp. 3449–3461
2019
Later among the works it cites.
D. Bau, H. Strobelt, W. Peebles, J. Wulff, B. Zhou, J. Zhu, A. Torralba, Semantic photo manipulation with a generative image prior, ACM Transactions on Graphics 38 (4) (2019)
2019
Later among the works it cites.
Y. Jia, R. J. Weiss, F. Biadsy, W. Macherey, M. Johnson, Z. Chen, Y. Wu, Direct speech-to-speech translation with a sequence-to-sequence model, in: INTERSPEECH, 2019
2019
Later among the works it cites.
L. Wang, W. Chen, W. Yang, F. Bi, F. R. Yu, A state-of-the-art review on image synthesis with generative adversarial networks, IEEE Access 8 (2020) 63514–63537
2020
Later among the works it cites.
J. Agnese, J. Herrera, H. Tao, X. Zhu, A survey and taxonomy of adversarial neural networks for text-to-image synthesis, Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
D. Pavllo, A. Lucchi, T. Hofmann, Controlling style and semantics in weakly-supervised image generation, in: European Conference on Computer Vision, 2020, pp. 482–499
2020
Later among the works it cites.
T.-Y. Lin, P. Goyal, R. B. Girshick, K. He, P. Dollár, Focal loss for dense object detection, IEEE Transactions on Pattern Analysis and Machine Intelligence 42 (2020) 318–327
2020
Later among the works it cites.
D. Stap, M. Bleeker, S. Ibrahimi, M. ter Hoeve, Conditional image generation and manipulation for user-specified content, in: Proceedings of the IEEE Computer Vision and Pattern Recognition Workshop, 2020
2020
Later among the works it cites.
Z. Wang, Z. Quan, Z. Wang, X. Hu, Y. Chen, Text to image synthesis with bidirectional generative adversarial network, in: IEEE International Conference on Multimedia and Expo, 2020, pp. 1–6
2020
Later among the works it cites.
R. Rombach, P. Esser, B. Ommer, Network-to-network translation with conditional invertible neural networks, Advances in Neural Information Processing Systems 33 (2020)
2020
Later among the works it cites.
J. Cheng, F. Wu, Y. Tian, L. Wang, D. Tao, Rifegan: Rich feature generation for text-to-image synthesis from prior knowledge, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2020, pp. 10911–10920
2020
Later among the works it cites.
T. Niu, F. Feng, L. Li, X. Wang, Image synthesis from locally related texts, Proceedings of the International Conference on Multimedia Retrieval (2020)
2020
Later among the works it cites.
S. Frolov, S. Jolly, J. Hees, A. Dengel, Leveraging visual question answering to improve text-to-image synthesis, in: Proceedings of the Second Workshop on Beyond Vision and LANguage: inTEgrating Real-world kNowledge, 2020, pp. 17–22
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
T. Hinz, S. Heinrich, S. Wermter, Semantic object accuracy for generative text-to-image synthesis, IEEE Transactions on Pattern Analysis and Machine Intelligence (2020)
2020
Later among the works it cites.
M. Wang, C. Lang, L. Liang, G. Lyu, S. Feng, T. Wang, Attentive generative adversarial network to bridge multi-domain gap for image synthesis, in: IEEE International Conference on Multimedia and Expo, 2020, pp. 1–6
2020
Later among the works it cites.
M. Wang, C. Lang, L. Liang, S. Feng, T. Wang, Y. Gao, End-to-end text-to-image synthesis with spatial constrains, ACM Transactions on Intelligent Systems and Technology 11 (4) (2020) 1–19
2020
Later among the works it cites.
Z. Wu, S. Pan, F. Chen, G. Long, C. Zhang, P. S. Yu, A comprehensive survey on graph neural networks, IEEE Transactions on Neural Networks and Learning Systems (2020)
2020
Later among the works it cites.
D. M. Vo, A. Sugimoto, Visual-relation conscious image generation from structured-text, in: European Conference on Computer Vision, Springer, 2020, pp. 290–306
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Pont-Tuset, J. Uijlings, S. Changpinyo, R. Soricut, V. Ferrari, Connecting vision and language with localized narratives, in: European Conference on Computer Vision, 2020, pp. 647–664
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, B. Poole, Score-based generative modeling through stochastic differential equations, 2011.13456 (2020)
2020
Later among the works it cites.
M. Chen, A. Radford, R. Child, J. Wu, H. Jun, D. Luan, I. Sutskever, Generative pretraining from pixels, in: International Conference on Machine Learning, 2020, pp. 1691–1703
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
D. Bau, S. Liu, T. Wang, J.-Y. Zhu, A. Torralba, Rewriting a deep generative model, in: European Conference on Computer Vision, 2020, pp. 351–369
2020
Later among the works it cites.
D. Zhu, A. Mogadala, D. Klakow, Image manipulation with natural language using two-sided attentive conditional generative adversarial network, Neural Networks (2020)
2020
Later among the works it cites.
Y. Liu, M. De Nadai, D. Cai, H. Li, X. Alameda-Pineda, N. Sebe, B. Lepri, Describe what to change: A text-guided unsupervised image-to-image translation approach, in: Proceedings of the ACM International Conference on Multimedia, 2020, pp. 1357–1365
2020
Later among the works it cites.
L. Zhang, Q. Chen, B. Hu, S. Jiang, Text-guided neural image inpainting, in: Proceedings of the ACM International Conference on Multimedia, 2020, pp. 1302–1310
2020
Later among the works it cites.
B. Li, X. Qi, T. Lukasiewicz, P. H. Torr, Manigan: Text-guided image manipulation, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2020, pp. 7880–7889
2020
Later among the works it cites.
H.-S. Choi, C.-D. Park, K. Lee, From inference to generation: End-to-end fully self-supervised generation of human face from speech, in: International Conference on Learning Representations, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Liang, W. Pei, F. Lu, Cpgan: Content-parsing generative adversarial networks for text-to-image synthesis, in: European Conference on Computer Vision, 2020, pp. 491–508
2020
Later among the works it cites.
2021
Closest in time.
12312, Dall·e: Creating images from text, https://openai.com/blog/dall-e/ , [Accessed 21-January-2021] (2021)
2021
Closest in time.
D. Suris, A. Recasens, D. Bau, D. Harwath, J. Glass, A. Torralba, Learning words by drawing images, in: Proceedings of the IEEE Computer Vision and Pattern Recognition, 2019, pp. 2029–2038
2038
Closest in time.
K. Xu, J. Ba, R. Kiros, K. Cho, A. C. Courville, R. Salakhutdinov, R. S. Zemel, Y. Bengio, Show, attend and tell: Neural image caption generation with visual attention, in: International Conference on Machine Learning, 2015, pp. 2048–2057
2057
Closest in time.