Fetching the paper…
Reading the bibliography…
Visual storytelling aims to generate compelling narratives from image sequences.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Logical necessity and transitivity of causal relations in stories
Tom Trabasso, Paul Van den Broek, and So Young Suh. 1989 · 1989
Earlier work this paper cites.
Issue-based dialogue management
Staffan Larsson. 2002 · 2002
Earlier work this paper cites.
Content planning for neural story generation with aristotelian rescoring
Seraphina Goldfarb-Tarrant, Tuhin Chakrabarty, Ralph Weischedel, and Nanyun Peng. 2020a · 2009
Earlier work this paper cites.
Information structure: Towards an integrated formal theory of pragmatics
Craige Roberts. 2012 · 2012
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Visual storytelling
Ting-Hao Huang, Francis Ferraro, Nasrin Mostafazadeh, Ishan Misra, Aishwarya Agrawal, Jacob Devlin, Ross Girshick, Xiaodong He, Pushmeet Kohli, Dhruv Batra, et al. 2016 · 2016
Earlier work this paper cites.
Visual relationship detection with language priors
Cewu Lu, Ranjay Krishna, Michael Bernstein, and Li Fei-Fei. 2016 · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016 · 2016
Earlier work this paper cites.
A simpler and more generalizable story detector using verb and character features
Joshua Eisenberg and Mark Finlayson. 2017 · 2017
Earlier work this paper cites.
Contextualize, show and tell: A neural visual storyteller
Diana Gonzalez-Rico and Gibran Fuentes-Pineda. 2018 · 2018
Earlier work this paper cites.
Glac net: Glocal attention cascading networks for multi-image cued story generation
Taehyeong Kim, Min-Oh Heo, Seonil Son, Kyoung-Wha Park, and Byoung-Tak Zhang. 2018 · 2018
Earlier work this paper cites.
Know what you don’t know: Unanswerable questions for squad
Pranav Rajpurkar, Robin Jia, and Percy Liang. 2018 · 2018
Earlier work this paper cites.
No metrics are perfect: Adversarial reward learning for visual storytelling
Xin Wang, Wenhu Chen, Yuan-Fang Wang, and William Yang Wang. 2018 · 2018
Earlier work this paper cites.
A skeleton-based model for promoting coherence among sentences in narrative story generation
Jingjing Xu, Xuancheng Ren, Yi Zhang, Qi Zeng, Xiaoyan Cai, and Xu Sun. 2018 · 2018
Earlier work this paper cites.
Synthetic QA corpora generation with roundtrip consistency
Chris Alberti, Daniel Andor, Emily Pitler, Jacob Devlin, and Michael Collins. 2019 · 2019
Earlier work this paper cites.
Strategies for structuring story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2019 · 2019
Earlier work this paper cites.
The steep road to happily ever after: an analysis of current visual storytelling models
Yatri Modi and Natalie Parde. 2019 · 2019
Cited alongside, same era.
Step-by-step: Separating planning from realization in neural data-to-text generation
Amit Moryossef, Yoav Goldberg, and Ido Dagan. 2019 · 2019
Cited alongside, same era.
Knowledgeable storyteller: A commonsense-driven generative model for visual storytelling
Pengcheng Yang, Fuli Luo, Peng Chen, Lei Li, Zhiyi Yin, Xiaodong He, and Xu Sun. 2019 · 2019
Cited alongside, same era.
Plan-and-write: Towards better automatic storytelling
Lili Yao, Nanyun Peng, Ralph Weischedel, Kevin Knight, Dongyan Zhao, and Rui Yan. 2019 · 2019
Cited alongside, same era.
Towards automatically generating questions under discussion to link information and discourse structure
Kordula De Kuthy, Madeeswaran Kannan, Haemanth Santhi Ponnusamy, and Detmar Meurers. 2020 · 2020
Cited alongside, same era.
Controllable neural dialogue summarization with personal named entity planning
Zhengyuan Liu and Nancy Chen. 2021 · 2021
Later among the works it cites.
Clipcap: Clip prefix for image captioning
Ron Mokady, Amir Hertz, and Amit H Bermano. 2021 · 2021
Later among the works it cites.
Planning with learned entity prompts for abstractive summarization
Shashi Narayan, Yao Zhao, Joshua Maynez, Gonçalo Simões, Vitaly Nikolaev, and Ryan McDonald. 2021 · 2021
Later among the works it cites.
Mauve: Measuring the gap between neural text and human text using divergence frontiers
Krishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun, Sean Welleck, Yejin Choi, and Zaid Harchaoui. 2021 · 2021
Later among the works it cites.
T5 (base) fine-tuned on squad for qg via ap
Manuel Romero. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Content planning for neural story generation with aristotelian rescoring
Seraphina Goldfarb-Tarrant, Tuhin Chakrabarty, Ralph Weischedel, and Nanyun Peng. 2020b · 2020
Cited alongside, same era.
Diverse and relevant visual storytelling with scene graph embeddings
Xudong Hong, Rakshith Shetty, Asad Sayeed, Khushboo Mehra, Vera Demberg, and Bernt Schiele. 2020 · 2020
Cited alongside, same era.
Knowledge-enriched visual storytelling
Chao-Chun Hsu, Zi-Yuan Chen, Chi-Yang Hsu, Chih-Chia Li, Tzu-Yuan Lin, Ting-Hao Huang, and Lun-Wei Ku. 2020 · 2020
Cited alongside, same era.
What makes a good story? designing composite rewards for visual storytelling
Junjie Hu, Yu Cheng, Zhe Gan, Jingjing Liu, Jianfeng Gao, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
A character-centric neural model for automated story generation
Danyang Liu, Juntao Li, Meng-Hsuan Yu, Ziming Huang, Gongshen Liu, Dongyan Zhao, and Rui Yan. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Cited alongside, same era.
Maria Tsimpoukelli, Jacob L Menick, Serkan Cabi, SM Eslami, Oriol Vinyals, and Felix Hill. 2021 · 2021
Later among the works it cites.
Imagine, reason and write: Visual storytelling with graph knowledge and relational reasoning
Chunpu Xu, Min Yang, Chengming Li, Ying Shen, Xiang Ao, and Ruifeng Xu. 2021 · 2021
Later among the works it cites.
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katherine Millican, Malcolm Reynolds, et al. 2022 · 2022
Later among the works it cites.
Qafacteval: Improved qa-based factual consistency evaluation for summarization
Alexander Richard Fabbri, Chien-Sheng Wu, Wenhao Liu, and Caiming Xiong. 2022 · 2022
Later among the works it cites.
Learning to rank visual stories from human ranking data
Chi-Yang Hsu, Yun-Wei Chu, Vincent Chen, Kuan-Chieh Lo, Chacha Chen, Ting-Hao Huang, and Lun-Wei Ku. 2022 · 2022
Later among the works it cites.
Conditional generation with a question-answering blueprint
Shashi Narayan, Joshua Maynez, Reinald Kim Amplayo, Kuzman Ganchev, Annie Louis, Fantine Huot, Dipanjan Das, and Mirella Lapata. 2022 · 2022
Later among the works it cites.
Data-to-text generation with variational sequential planning
Ratish Puduppully, Yao Fu, and Mirella Lapata. 2022 · 2022
Later among the works it cites.
Rovist: Learning robust metrics for visual storytelling
Eileen Wang, Caren Han, and Josiah Poon. 2022 · 2022
Later among the works it cites.
Re3: Generating longer stories with recursive reprompting and revision
Kevin Yang, Nanyun Peng, Yuandong Tian, and Dan Klein. 2022 · 2022
Later among the works it cites.
Lit: Zero-shot transfer with locked-image text tuning
Xiaohua Zhai, Xiao Wang, Basil Mustafa, Andreas Steiner, Daniel Keysers, Alexander Kolesnikov, and Lucas Beyer. 2022 · 2022
Later among the works it cites.
Incorporating question answering-based signals into abstractive summarization via salient span selection
Daniel Deutsch and Dan Roth. 2023 · 2023
Closest in time.
Language is not all you need: Aligning perception with language models
Shaohan Huang, Li Dong, Wenhui Wang, Yaru Hao, Saksham Singhal, Shuming Ma, Tengchao Lv, Lei Cui, Owais Khan Mohammed, Qiang Liu, et al. 2023 · 2023
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2023 · 2023
Closest in time.