Fetching the paper…
Reading the bibliography…
We propose a new task, called Story Visualization.
Image quality metrics: Psnr vs. ssim
A. Hore and D. Ziou · 2010
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Nice: Non-linear independent components estimation
L. Dinh, D. Krueger, and Y. Bengio · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Density estimation using real nvp
L. Dinh, J. Sohl-Dickstein, and S. Bengio · 2016
Earlier work this paper cites.
Visual storytelling
T.-H. K. Huang, F. Ferraro, N. Mostafazadeh, I. Misra, A. Agrawal, J. Devlin, R. Girshick, X. He, P. Kohli, D. Batra, et al · 2016
Earlier work this paper cites.
Dynamic filter networks
X. Jia, B. De Brabandere, T. Tuytelaars, and L. V. Gool · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
S. Reed, Z. Akata, X. Yan, L. Logeswaran, B. Schiele, and H. Lee · 2016
Earlier work this paper cites.
Generating videos with scene dynamics
C. Vondrick, H. Pirsiavash, and A. Torralba · 2016
Earlier work this paper cites.
Attribute2image: Conditional image generation from visual attributes
X. Yan, J. Yang, K. Sohn, and H. Lee · 2016
Earlier work this paper cites.
Unsupervised learning of disentangled representations from video
E. L. Denton et al · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros · 2017
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Cited alongside, same era.
Codraw: Visual dialog for collaborative drawing
J.-H. Kim, D. Parikh, D. Batra, B.-T. Zhang, and Y. Tian · 2017
Cited alongside, same era.
Deepstory: Video story qa by deep embedded memory networks
K.-M. Kim, M.-O. Heo, S.-H. Choi, and B.-T. Zhang · 2017
Cited alongside, same era.
Recurrent topic-transition gan for visual paragraph generation
X. Liang, Z. Hu, H. Zhang, C. Gan, and E. P. Xing · 2017
Cited alongside, same era.
Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks
H. Zhang, T. Xu, H. Li, S. Zhang, X. Huang, X. Wang, and D. Metaxas · 2017
Cited alongside, same era.
Controllable video generation with sparse trajectories
Z. Hao, X. Huang, and S. Belongie · 2018
Closest in time.
Probabilistic video generation using holistic attribute control
J. He, A. Lehrmann, J. Marino, G. Mori, and L. Sigal · 2018
Closest in time.
Hierarchically structured reinforcement learning for topically coherent visual story generation
Q. Huang, Z. Gan, A. Celikyilmaz, D. Wu, J. Wang, and X. He · 2018
Closest in time.
Video generation from text
Y. Li, M. R. Min, D. Shen, D. Carlson, and L. Carin · 2018
Closest in time.
Show me a story: Towards coherent neural story illustration
H. Ravi, L. Wang, C. Muniz, L. Sigal, D. Metaxas, and M. Kapadia · 2018
Closest in time.
Efficient parametrization of multi-domain deep neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unpaired image-to-image translation using cycle-consistent adversarial networks
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Cited alongside, same era.
Deep video generation, prediction and completion of human action sequences
H. Cai, C. Bai, Y.-W. Tai, and C.-K. Tang · 2018
Cited alongside, same era.
D. Cer, Y. Yang, S.-y. Kong, N. Hua, N. Limtiaco, R. S. John, N. Constant, M. Guajardo-Cespedes, S. Yuan, C. Tar, et al · 2018
Cited alongside, same era.
Language-based image editing with recurrent attentive models
J. Chen, Y. Shen, J. Gao, J. Liu, and X. Liu · 2018
Cited alongside, same era.
Sequential attention gan for interactive image editing via dialogue
Y. Cheng, Z. Gan, Y. Li, J. Liu, and J. Gao · 2018
Cited alongside, same era.
Stochastic video generation with a learned prior
E. Denton and R. Fergus · 2018
Cited alongside, same era.
Keep drawing it: Iterative language-based image generation and editing
A. El-Nouby, S. Sharma, H. Schulz, D. Hjelm, L. E. Asri, S. E. Kahou, Y. Bengio, and G. W. Taylor · 2018
Cited alongside, same era.
S.-A. Rebuffi, H. Bilen, and A. Vedaldi · 2018
Closest in time.
Chatpainter: Improving text to image generation using dialogue
S. Sharma, D. Suhubdy, V. Michalski, S. E. Kahou, and Y. Bengio · 2018
Closest in time.
Adversarial scene editing: Automatic object removal from weak supervision
R. Shetty, M. Fritz, and B. Schiele · 2018
Closest in time.
Mocogan: Decomposing motion and content for video generation
S. Tulyakov, M.-Y. Liu, X. Yang, and J. Kautz · 2018
Closest in time.
Every smile is unique: Landmark-guided diverse smile generation
W. Wang, X. Alameda-Pineda, D. Xu, P. Fua, E. Ricci, and N. Sebe · 2018
Closest in time.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
T. Xu, P. Zhang, Q. Huang, H. Zhang, Z. Gan, X. Huang, and X. He · 2018
Closest in time.
Learning to forecast and refine residual motion for image-to-video generation
L. Zhao, X. Peng, Y. Tian, M. Kapadia, and D. Metaxas · 2018
Closest in time.