Fetching the paper…
Reading the bibliography…
We show how we can globally edit images using textual instructions: given a source image and a textual instruction for the edit, generate a new image transformed under this instruction.
Markov random fields and their applications
Kindermann, R., Snell, J.L.: · 1980
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, V., Hinton, G.E.: · 2010
Earlier work this paper cites.
Learning photographic global tonal adjustment with a database of input / output image pairs
Bychkovsky, V., Paris, S., Chan, E., Durand, F.: · 2011
Earlier work this paper cites.
Context-based automatic local image enhancement
Hwang, S., Kapoor, A., Kang, S.B.: · 2012
Earlier work this paper cites.
Pixeltone: A multimodal interface for image editing
Laput, G., Dontcheva, M., Wilensky, G., Chang, W., Agarwala, A., Linder, J., Adar, E.: · 2013
Earlier work this paper cites.
Collaborative personalization of image enhancement
Kapoor, A., Caicedo, J.C., Lischinski, D., Kang, S.B.: · 2013
Earlier work this paper cites.
Grounded compositional semantics for finding and describing images with sentences
Socher, R., Karpathy, A., Le, Q.V., Manning, C.D., Ng, A.Y.: · 2013
Earlier work this paper cites.
Learning deep structured semantic models for web search using clickthrough data
Huang, P.S., He, X., Gao, J., Deng, L., Acero, A., Heck, L.: · 2013
Earlier work this paper cites.
Imagespirit: Verbal guided image parsing
Cheng, M.M., Zheng, S., Lin, W.Y., Vineet, V., Sturgess, P., Croo, N., Mitra, N., Torr, P.: · 2014
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, I.J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., Bengio, Y.: · 2014
Earlier work this paper cites.
A learning-to-rank approach for image color enhancement
Yan, J., Lin, S., Kang, S.B., Tang, X.: · 2014
Earlier work this paper cites.
Color transfer using probabilistic moving least squares
Hwang, Y., Lee, J.Y., Kweon, I.S., Kim, S.J.: · 2014
Earlier work this paper cites.
Autostyle: Automatic style transfer from image collections to users’ images
Liu, Y., Cohen, M., Uyttendaele, M., Rusinkiewicz, S.: · 2014
Earlier work this paper cites.
Deep fragment embeddings for bidirectional image sentence mapping
Karpathy, A., Joulin, A., Li, F.F.: · 2014
Earlier work this paper cites.
What are you talking about? text-to-image coreference
Kong, C., Lin, D., Bansal, M., Urtasun, R., Fidler, S.: · 2014
Earlier work this paper cites.
The stanford corenlp natural language processing toolkit
Manning, C., Surdeanu, M., Bauer, J., Finkel, J., Bethard, S., McClosky, D.: · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
Mirza, M., Osindero, S.: · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K., Zisserman, A.: · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.Y., Maire, M., Belongie, S., Bourdev, L., Girshick, R., Hays, J., Perona, P., Ramanan, D., Zitnick, C.L., Dollár, P.: · 2014
Earlier work this paper cites.
Referit game: Referring to objects in photographs of natural scenes
Kazemzadeh, S., Ordonez, V., Matten, M., Berg, T.L.: · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D.P., Ba, J.: · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, J., Gulcehre, C., Cho, K.H., Bengio, Y.: · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Pennington, J., Socher, R., Manning, C.D.: · 2014
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Luong, T., Pham, H., Manning, C.D.: · 2015
Cited alongside, same era.
Automatic photo adjustment using deep neural networks
Yan, Z., Zhang, H., Wang, B., Paris, S., Yu, Y.: · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
Vinyals, O., Toshev, A., Bengio, S., Erhan, D.: · 2015
Cited alongside, same era.
Mind’s eye: A recurrent visual representation for image caption generation
Chen, X., Zitnick, C.L.: · 2015
Cited alongside, same era.
Generative adversarial text-to-image synthesis
Reed, S., Akata, Z., Yan, X., Logeswaran, L., Schiele, B., Lee, H.: · 2016
Later among the works it cites.
Attribute2image: Conditional image generation from visual attributes
Yan, X., Yang, J., Sohn, K., Lee, H.: · 2016
Later among the works it cites.
Generation and comprehension of unambiguous object descriptions
Mao, J.: · 2016
Later among the works it cites.
Instance normalization: The missing ingredient for fast stylization
Ulyanov, D., Vedaldi, A., Lempitsky, V.: · 2016
Later among the works it cites.
Perceptual losses for real-time style transfer and super-resolution
Johnson, J., Alahi, A., Fei-Fei, L.: · 2016
Later among the works it cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., Sun, J.: · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Long-term recurrent convolutional networks for visual recognition and description
Donahue, J., Darrell, T.: · 2015
Cited alongside, same era.
Translating videos to natural language using deep recurrent neural networks
Venugopalan, S., Xu, H., Saenko, K.: · 2015
Cited alongside, same era.
Vqa: Visual question answering
Antol, S., Agrawal, A., Batra, D., Zitnick, C.L., Parikh, D.: · 2015
Cited alongside, same era.
Neural self talk: Image understanding via continuous questioning and answering
Yang, Y., Li, Y., Fermuller, C., Aloimonos, Y.: · 2015
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
Karpathy, A., Li, F.F.: · 2015
Cited alongside, same era.
Improved semantic representations from tree-structured long short-term memory networks
Tai, K.S., Socher, R., Manning, C.D.: · 2015
Cited alongside, same era.
Flickr30k entities: Collecting region-to-phrase correspondences for richer image-to-sentence models
Plummer, B.A.: · 2016
Later among the works it cites.
Language-based image editing with recurrent attentive models
Chen, J., Shen, Y., Gao, J., Liu, J., Liu, X.: · 2017
Later among the works it cites.
Cross-sentence n-ary relation extraction with graph lstms
Peng, N., Poon, H., Quirk, C., Toutanova, K., Yih, W.t.: · 2017
Later among the works it cites.
Towards diverse and natural image descriptions via a conditional gan
Dai, B., Lin, D., Urtasun, R., Fidler, S.: · 2017
Later among the works it cites.
Visual dialog
Das, A., Kottur, S., Gupta, K., Singh, A., Yadav, D., Moura, J.M.F., Parikh, D., Batra, D.: · 2017
Later among the works it cites.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Xu, T., Zhang, P., Huang, Q., Zhang, H., Gan, Z., Huang, X., He, X.: · 2017
Later among the works it cites.
A joint speaker-listener-reinforcer model for referring expressions
Yu, L., Tan, H., Bansal, M., Berg, T.L.: · 2017
Later among the works it cites.
Comprehension-guided referring expressions
Luo, R., Shakhnarovich, G.: · 2017
Later among the works it cites.
Photo-realistic single image super-resolution using a generative adversarial network
Ledig, C., Theis, L., Huszar, F., Caballero, J., Cunningham, A., Acosta, A., Aitken, A., Tejani, A., Totz, J., Wang, Z., Shi, W.: · 2017
Later among the works it cites.
Stylebank: An explicit representation for neural image style transfer
Chen, D., Yuan, L., Liao, J., Yu, N., Hua, G.: · 2017
Later among the works it cites.
Image-to-image translation with conditional adversarial networks
Isola, P., Zhu, J.Y., Zhou, T., Efros, A.A.: · 2017
Later among the works it cites.
I2t2i: Learning text to image synthesis with textual data augmentation
Dong, H., Zhang, J., McIlwraith, D., Guo, Y.: · 2017
Later among the works it cites.
Interactive image manipulation with natural language instruction commands
Seitaro, S., Koichiro, Y., Sakriani, S., Yu, S., Satoshi, N.: · 2018
Closest in time.
Inferring semantic layout for hierarchical text-to-image synthesis
Hong, S., Yang, D., Choi, J., Lee, H.: · 2018
Closest in time.
Chatpainter: Improving text to image generation using dialogue
Sharma, S., Suhubdy, D., Michalski, V., Kahou, S.E., Bengio, Y.: · 2018
Closest in time.