Fetching the paper…
Reading the bibliography…
Visual storytelling systems generate multi-sentence stories from image sequences.
Coherence and structure in text and discourse
Gisela Redeker. 2000 · 2000
Earlier work this paper cites.
Narrative prose generation
Charles B. Callaway and James C. Lester. 2001 · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Automatic evaluation of machine translation quality using longest common subsequence and skip-bigram statistics
Chin-Yew Lin and Franz Josef Och. 2004 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Comparing Automatic and Human Evaluation of NLG Systems
Anja Belz and E Reiter. 2006 · 2006
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston. 2009 · 2009
Earlier work this paper cites.
Computational approaches to storytelling and creativity
Pablo Gervás. 2009 · 2009
Earlier work this paper cites.
An Investigation into the Validity of Some Metrics for Automatically Evaluating Natural Language Generation Systems
Ehud Reiter and Anja Belz. 2009 · 2009
Earlier work this paper cites.
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
SPICE: semantic propositional image caption evaluation
Peter Anderson, Basura Fernando, Mark Johnson, and Stephen Gould. 2016 · 2016
Earlier work this paper cites.
Visual storytelling
Ting-Hao Kenneth Huang, Francis Ferraro, Nasrin Mostafazadeh, Ishan Misra, Aishwarya Agrawal, Jacob Devlin, Ross Girshick, Xiaodong He, Pushmeet Kohli, Dhruv Batra, C Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Michel Galley, and Margaret Mitchell. 2016 · 2016
Earlier work this paper cites.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2016 · 2016
Earlier work this paper cites.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Earlier work this paper cites.
Contextualize, show and tell: A neural visual storyteller
Diana Gonzalez-Rico and Gibran Fuentes-Pineda. 2018 · 2018
Earlier work this paper cites.
GLAC net: GLocal attention cascading networks for multi-image cued story generation
Taehyeong Kim, Min-Oh Heo, Seonil Son, Kyoung-Wha Park, and Byoung-Tak Zhang. 2018 · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Earlier work this paper cites.
A Structured Review of the Validity of BLEU
Ehud Reiter. 2018 · 2018
Earlier work this paper cites.
“my way of telling a story”: Persona based grounded story generation
Khyathi Chandu, Shrimai Prabhumoye, Ruslan Salakhutdinov, and Alan W Black. 2019 · 2019
Earlier work this paper cites.
Hierarchically structured reinforcement learning for topically coherent visual story generation
Qiuyuan Huang, Zhe Gan, Asli Celikyilmaz, Dapeng Wu, Jianfeng Wang, and Xiaodong He. 2019 · 2019
Earlier work this paper cites.
Emotion reinforced visual storytelling
Nanxing Li, Bei Liu, Zhizhong Han, Yu-Shen Liu, and Jianlong Fu. 2019b · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
Knowledgeable storyteller: A commonsense-driven generative model for visual storytelling
Pengcheng Yang, Fuli Luo, Peng Chen, Lei Li, Zhiyi Yin, Xiaodong He, and Xu Sun. 2019 · 2019
Cited alongside, same era.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Knowledge-enriched visual storytelling
Chao-Chun Hsu, Zi-Yuan Chen, Chi-Yang Hsu, Chih-Chia Li, Tzu-Yuan Lin, Ting-Hao Kenneth Huang, and Lun-Wei Ku. 2020 · 2020
Cited alongside, same era.
What makes a good story? designing composite rewards for visual storytelling
Junjie Hu, Yu Cheng, Zhe Gan, Jingjing Liu, Jianfeng Gao, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
How Can We Know What Language Models Know?
PADA: Example-based Prompt Learning for on-the-fly Adaptation to Unseen Domains
Eyal Ben-David, Nadav Oved, and Roi Reichart. 2022 · 2022
Later among the works it cites.
Visual storytelling with hierarchical BERT semantic guidance
Ruichao Fan, Hanli Wang, Jinjing Gu, and Xianhui Liu. 2022 · 2022
Later among the works it cites.
Knowledge-enriched attention network with group-wise semantic for visual storytelling
Tengpeng Li, Hanli Wang, Bin He, and Chang Wen Chen. 2022 · 2022
Later among the works it cites.
Human Evaluation and Correlation with Automatic Metrics in Consultation Note Generation
Francesco Moramarco, Alex Papadopoulos Korfiatis, Mark Perera, Damir Juric, Jack Flann, Ehud Reiter, Anya Belz, and Aleksandar Savkov. 2022 · 2022
Later among the works it cites.
A contrastive framework for neural text generation
Yixuan Su, Tian Lan, Yan Wang, Dani Yogatama, Lingpeng Kong, and Nigel Collier. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
BLEURT: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, and Ankur Parikh. 2020 · 2020
Cited alongside, same era.
Storytelling from an image stream using scene graphs
Ruize Wang, Zhongyu Wei, Piji Li, Qi Zhang, and Xuanjing Huang. 2020 · 2020
Cited alongside, same era.
Commonsense knowledge aware concept selection for diverse and informative visual storytelling
Hong Chen, Yifei Huang, Hiroya Takamura, and Hideki Nakayama. 2021 · 2021
Cited alongside, same era.
BERTese: Learning to speak to BERT
Adi Haviv, Jonathan Berant, and Amir Globerson. 2021 · 2021
Cited alongside, same era.
CLIPScore: A reference-free evaluation metric for image captioning
Jack Hessel, Ari Holtzman, Maxwell Forbes, Ronan Le Bras, and Yejin Choi. 2021 · 2021
Cited alongside, same era.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Cited alongside, same era.
RoViST: Learning robust metrics for visual storytelling
Eileen Wang, Caren Han, and Josiah Poon. 2022 · 2022
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90% chatgpt quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E. Gonzalez, Ion Stoica, and Eric P. Xing. 2023 · 2023
Later among the works it cites.
Autoad: Movie description in context
Tengda Han, Max Bain, Arsha Nagrani, Gül Varol, Weidi Xie, and Andrew Zisserman. 2023 · 2023
Later among the works it cites.
Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences
Xudong Hong, Asad Sayeed, Khushboo Mehra, Vera Demberg, and Bernt Schiele. 2023 · 2023
Later among the works it cites.
Grounding language models to images for multimodal inputs and outputs
Jing Yu Koh, Ruslan Salakhutdinov, and Daniel Fried. 2023 · 2023
Later among the works it cites.
Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi. 2023 · 2023
Later among the works it cites.
Detecting and grounding important characters in visual stories
Danyang Liu and Frank Keller. 2023 · 2023
Later among the works it cites.
Linearly mapping from image to text space
Jack Merullo, Louis Castricato, Carsten Eickhoff, and Ellie Pavlick. 2023 · 2023
Later among the works it cites.
GROOViST: A metric for grounding objects in visual storytelling
Aditya Surikuchi, Sandro Pezzelle, and Raquel Fernández. 2023 · 2023
Later among the works it cites.
Text-only training for visual storytelling
Yuechen Wang, Wengang Zhou, Zhenbo Lu, and Houqiang Li. 2023 · 2023
Later among the works it cites.
Attractive storyteller: Stylized visual storytelling with unpaired text
Dingyi Yang and Qin Jin. 2023 · 2023
Later among the works it cites.
Visualize before you write: Imagination-guided open-ended text generation
Wanrong Zhu, An Yan, Yujie Lu, Wenda Xu, Xin Wang, Miguel Eckstein, and William Yang Wang. 2023 · 2023
Later among the works it cites.
Introducing Meta Llama 3: The most capable openly available LLM to date
Meta AI. 2024 · 2024
Closest in time.
Mistral llms
Mistral AI. 2024 · 2024
Closest in time.
When do prompting and prefix-tuning work? a theory of capabilities and limitations
Aleksandar Petrov, Philip Torr, and Adel Bibi. 2024 · 2024
Closest in time.
SCO-VIST: Social interaction commonsense knowledge-based visual storytelling
Eileen Wang, Caren Han, and Josiah Poon. 2024 · 2024
Closest in time.