Fetching the paper…
Reading the bibliography…
Most existing text-to-image synthesis tasks are static single-turn generation, based on pre-defined textual descriptions of images.
Discriminative learning in sequential pattern recognition
X. He, L. Deng, and W. Chou · 2008
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Amazon’s mechanical turk: A new source of inexpensive, yet high-quality, data?
M. Buhrmester, T. Kwang, and S. Gosling · 2011
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
M. Mirza and S. Osindero · 2014
Earlier work this paper cites.
Fine-grained visual comparisons with local learning
A. Yu and K. Grauman · 2014
Earlier work this paper cites.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Segmentation from natural language expressions
R. Hu, M. Rohrbach, and T. Darrell · 2016
Earlier work this paper cites.
Deepfashion: Powering robust clothes recognition and retrieval with rich annotations
Z. Liu, P. Luo, S. Qiu, X. Wang, and X. Tang · 2016
Earlier work this paper cites.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krähenbühl, J. Donahue, T. Darrell, and A. Efros · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
S. Reed, Z. Akata, X. Yan, L. Logeswaran, B. Schiele, and H. Lee · 2016
Earlier work this paper cites.
Grounding of textual phrases in images by reconstruction
A. Rohrbach, M. Rohrbach, R. Hu, T. Darrell, and B. Schiele · 2016
Earlier work this paper cites.
Building end-to-end dialogue systems using generative hierarchical neural network models
I. V. Serban, A. Sordoni, Y. Bengio, A. Courville, and J. Pineau · 2016
Earlier work this paper cites.
Learning deep structure-preserving image-text embeddings
L. Wang, Y. Li, and S. Lazebnik · 2016
Earlier work this paper cites.
Generative image modeling using style and structure adversarial networks
X. Wang and A. Gupta · 2016
Cited alongside, same era.
Learning end-to-end goal-oriented dialog
A. Bordes, Y.-L. Boureau, and J. Weston · 2017
Cited alongside, same era.
Visual dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra · 2017
Cited alongside, same era.
Guesswhat?! visual object discovery through multi-modal dialogue
H. de Vries, F. Strub, S. Chandar, O. Pietquin, H. Larochelle, and A. C. Courville · 2017
Cited alongside, same era.
Aga: Attribute guided augmentation
M. Dixit, R. Kwitt, M. Niethammer, and N. Vasconcelos · 2017
Cited alongside, same era.
Image-to-image translation with conditional adversarial networks
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros · 2017
Cited alongside, same era.
Unpaired image-to-image translation using cycle-consistent adversarial networks
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Later among the works it cites.
Toward multimodal image-to-image translation
J.-Y. Zhu, R. Zhang, D. Pathak, T. Darrell, A. A. Efros, O. Wang, and E. Shechtman · 2017
Later among the works it cites.
Be your own prada: Fashion synthesis with structural coherence
S. Zhu, S. Fidler, R. Urtasun, D. Lin, and C. C. Loy · 2017
Later among the works it cites.
The neural painter: Multi-turn image generation
R. Y. Benmalek, C. Cardie, S. J. Belongie, X. He, and J. Gao · 2018
Closest in time.
Language-based image editing with recurrent attentive models
J. Chen, Y. Shen, J. Gao, J. Liu, and X. Liu · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Kim, D. Parikh, D. Batra, B. Zhang, and Y. Tian · 2017
Cited alongside, same era.
Mmd gan: Towards deeper understanding of moment matching network
C.-L. Li, W.-C. Chang, Y. Cheng, Y. Yang, and B. Poczos · 2017
Cited alongside, same era.
Unsupervised image-to-image translation networks
M.-Y. Liu, T. Breuel, and J. Kautz · 2017
Cited alongside, same era.
Using Reinforcement Learning to Model Incrementality in a Fast-Paced Dialogue Game
R. Manuvinakurike, D. DeVault, and K. Georgila · 2017
Cited alongside, same era.
Image-grounded conversations: Multimodal context for natural question and response generation
N. Mostafazadeh, C. Brockett, B. Dolan, M. Galley, J. Gao, G. Spithourakis, and L. Vanderwende · 2017
Cited alongside, same era.
Conditional image synthesis with auxiliary classifier GANs
A. Odena, C. Olah, and J. Shlens · 2017
Cited alongside, same era.
J. Gao, M. Galley, and L. Li · 2018
Closest in time.
Dialog-based interactive image retrieval
X. Guo, H. Wu, Y. Cheng, S. Rennie, and R. S. Feris · 2018
Closest in time.
Image generation from scene graphs
J. Johnson, A. Gupta, and L. Fei-Fei · 2018
Closest in time.
Conversational Image Editing: Incremental Intent Identification in a New Dialogue Task
R. Manuvinakurike, T. Bui, W. Chang, and K. Georgila · 2018
Closest in time.
Text-adaptive generative adversarial networks: Manipulating images with natural language
S. Nam, Y. Kim, and S. J. Kim · 2018
Closest in time.
Chatpainter: Improving text to image generation using dialogue
S. Sharma, D. Suhubdy, V. Michalski, S. E. Kahou, and Y. Bengio · 2018
Closest in time.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
T. Xu, P. Zhang, Q. Huang, H. Zhang, Z. Gan, X. Huang, and X. He · 2018
Closest in time.
Specifying object attributes and relations in interactive scene generation
O. Ashual and L. Wolf · 2019
Closest in time.
Storygan: A sequential conditional gan for story visualization
Y. Li, Z. Gan, Y. Shen, J. Liu, Y. Cheng, Y. Wu, L. Carin, D. Carlson, and J. Gao · 2019
Closest in time.
Image generation from layout
B. Zhao, L. Meng, W. Yin, and L. Sigal · 2019
Closest in time.
Bachgan: High-resolution image synthesis from salient object layout
Y. Li, Y. Cheng, Z. Gan, L. Yu, L. Wang, and J. Liu · 2020
Closest in time.