Fetching the paper…
Reading the bibliography…
Digital storytelling, essential in entertainment, education, and marketing, faces challenges in production scalability and flexibility.
Generation of small groups with rich behaviors from natural language interface
Chien-Yuan Chen, Sai-Keung Wong, and Wen-Yun Liu · 1960
Earlier work this paper cites.
Madame bovary on the holodeck: immersive interactive storytelling
Marc Cavazza, Jean-Luc Lugrin, David Pizzi, and Fred Charles · 2007
Earlier work this paper cites.
Toward supporting stories with procedurally generated game worlds
Ken Hartsook, Alexander Zook, Sauvik Das, and Mark O. Riedl · 2011
Earlier work this paper cites.
Digital storytelling: Capturing lives, creating community
Joe Lambert · 2013
Earlier work this paper cites.
How immersive is enough? a meta-analysis of the effect of immersive technology on user presence
James J. Cummings and Jeremy N. Bailenson · 2015
Earlier work this paper cites.
A survey on story generation techniques for authoring computational narratives
Ben Kybartas and Rafael Bidarra · 2016
Earlier work this paper cites.
Digital storytelling in research: A systematic review
Adele De Jager, Andrea Fogarty, Anna Tewson, Caroline Lenette, and Katherine M Boydell · 2017
Earlier work this paper cites.
Automated staging for virtual cinematography
Amaury Louarn, Marc Christie, and Fabrice Lamarche · 2018
Earlier work this paper cites.
Cardinal: Computer assisted authoring of movie scripts
Marcel Marti, Jodok Vieli, Wojciech Witoń, Rushit Sanghrajka, Daniel Inversini, Diana Wotruba, Isabel Simo, Sasha Schriber, Mubbasir Kapadia, and Markus Gross · 2018
Earlier work this paper cites.
A systematic review of educational digital storytelling
Jing Wu and Der-Thanq Victor Chen · 2019
Earlier work this paper cites.
The role of sound in inducing storytelling in immersive environments
Inês Salselas and Rui Penha · 2019
Earlier work this paper cites.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito · 2019
Earlier work this paper cites.
Stories of the town: balancing character autonomy and coherent narrative in procedurally generated worlds
Chris Miller, Mayank Dighe, Chris Martens, and Arnav Jhala · 2019
Earlier work this paper cites.
Story realization: Expanding plot events into sentences
Prithviraj Ammanabrolu, Ethan Tien, Wesley Cheung, Zhaochen Luo, William Ma, Lara J. Martin, and Mark O. Riedl · 2020
Earlier work this paper cites.
Controllable multi-character psychology-oriented story generation
Feifei Xu, Xinpeng Wang, Yunpu Ma, Volker Tresp, Yuyi Wang, Shanlin Zhou, and Haizhou Du · 2020
Earlier work this paper cites.
Write-an-animation: High-level text-based animation editing with character-scene interaction
Jia-Qi Zhang, Xiang Xu, Zhi-Meng Shen, Ze-Huan Huang, Yang Zhao, Yan-Pei Cao, Pengfei Wan, and Miao Wang · 2021
Earlier work this paper cites.
Segformer: Simple and efficient design for semantic segmentation with transformers
Enze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar, Jose M Alvarez, and Ping Luo · 2021
Earlier work this paper cites.
Imagen video: High definition video generation with diffusion models, 2022
Jonathan Ho, William Chan, Chitwan Saharia, Jay Whang, Ruiqi Gao, Alexey Gritsenko, Diederik P. Kingma, Ben Poole, Mohammad Norouzi, David J. Fleet, and Tim Salimans · 2022
Cited alongside, same era.
The ivi lab entry to the genea challenge 2022–a tacotron2 based method for co-speech gesture generation with locality-constraint attention mechanism
Che-Jui Chang, Sen Zhang, and Mubbasir Kapadia · 2022
Cited alongside, same era.
Procedural generation of narrative worlds
J. Timothy Balint and Rafael Bidarra · 2022
Cited alongside, same era.
Persona-guided planning for controlling the protagonist’s persona in story generation, 2022
Zhexin Zhang, Jiaxin Wen, Jian Guan, and Minlie Huang · 2022
Cited alongside, same era.
Storydall-e: Adapting pretrained text-to-image transformers for story continuation
Adyasha Maharana, Darryl Hannan, and Mohit Bansal · 2022
Cited alongside, same era.
Synthesizing physical character-scene interactions
Mohamed Hassan, Yunrong Guo, Tingwu Wang, Michael Black, Sanja Fidler, and Xue Bin Peng · 2023
Later among the works it cites.
The five-dollar model: generating game maps and sprites from sentence embeddings
Timothy Merino, Roman Negri, Dipika Rajesh, M Charity, and Julian Togelius · 2023
Later among the works it cites.
Storygpt-v: Large language models as consistent story visualizers, 2023
Xiaoqian Shen and Mohamed Elhoseiny · 2023
Later among the works it cites.
Story-to-motion: Synthesizing infinite and controllable character animation from long text
Zhongfei Qing, Zhongang Cai, Zhitao Yang, and Lei Yang · 2023
Later among the works it cites.
React: Synergizing reasoning and acting in language models, 2023
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2023
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models, 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improving image generation with better captions, 2023
James Betker, Gabriel Goh, Li Jing, Tim Brooks, Jianfeng Wang, Linjie Li, Long Ouyang, Juntang Zhuang, Joyce Lee, Yufei Guo, Wesam Manassra, Prafulla Dhariwal, Casey Chu, Yunxin Jiao, and Aditya Ramesh · 2023
Cited alongside, same era.
Wavjourney: Compositional audio creation with large language models, 2023
Xubo Liu, Zhongkai Zhu, Haohe Liu, Yi Yuan, Meng Cui, Qiushi Huang, Jinhua Liang, Yin Cao, Qiuqiang Kong, Mark D. Plumbley, and Wenwu Wang · 2023
Cited alongside, same era.
Audiogen: Textually guided audio generation, 2023
Felix Kreuk, Gabriel Synnaeve, Adam Polyak, Uriel Singer, Alexandre Défossez, Jade Copet, Devi Parikh, Yaniv Taigman, and Yossi Adi · 2023
Cited alongside, same era.
The importance of multimodal emotion conditioning and affect consistency for embodied conversational agents
Che-Jui Chang, Samuel S Sohn, Sen Zhang, Rajath Jayashankar, Muhammad Usman, and Mubbasir Kapadia · 2023
Cited alongside, same era.
Dreamoving: A human video generation framework based on diffusion models, 2023
Mengyang Feng, Jinlin Liu, Kai Yu, Yuan Yao, Zheng Hui, Xiefan Guo, Xianhui Lin, Haolan Xue, Chen Shi, Xiaowen Li, Aojie Li, Xiaoyang Kang, Biwen Lei, Miaomiao Cui, Peiran Ren, and Xuansong Xie · 2023
Cited alongside, same era.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning, 2023
Yuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang, Yaohui Wang, Yu Qiao, Maneesh Agrawala, Dahua Lin, and Bo Dai · 2023
Cited alongside, same era.
Magicedit: High-fidelity and temporally coherent video editing, 2023
Jun Hao Liew, Hanshu Yan, Jianfeng Zhang, Zhongcong Xu, and Jiashi Feng · 2023
Cited alongside, same era.
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior, 2023
Joon Sung Park, Joseph C. O’Brien, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2023
Later among the works it cites.
Communicative agents for software development, 2023
Chen Qian, Xin Cong, Wei Liu, Cheng Yang, Weize Chen, Yusheng Su, Yufan Dang, Jiahao Li, Juyuan Xu, Dahai Li, Zhiyuan Liu, and Maosong Sun · 2023
Later among the works it cites.
URL https://github.com/elevenlabs/elevenlabs-python
elevenlabs/elevenlabs-python, April 2024 · 2023
Later among the works it cites.
Playground v2.5: Three insights towards enhancing aesthetic quality in text-to-image generation, 2024
Daiqing Li, Aleks Kamko, Ehsan Akhgari, Ali Sabet, Linmiao Xu, and Suhail Doshi · 2024
Closest in time.
Scaling rectified flow transformers for high-resolution image synthesis, 2024
Patrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari, Jonas Müller, Harry Saini, Yam Levi, Dominik Lorenz, Axel Sauer, Frederic Boesel, Dustin Podell, Tim Dockhorn, Zion English, Kyle Lacey, Alex Goodwin, Yannik Marek, and Robin Rombach · 2024
Closest in time.
Video generation models as world simulators
Tim Brooks, Bill Peebles, Connor Holmes, Will DePue, Yufei Guo, Li Jing, David Schnurr, Joe Taylor, Troy Luhman, Eric Luhman, Clarence Ng, Ricky Wang, and Aditya Ramesh · 2024
Closest in time.
Generating human interaction motions in scenes with text control, 2024
Hongwei Yi, Justus Thies, Michael J. Black, Xue Bin Peng, and Davis Rempe · 2024
Closest in time.
Intelligent grimm – open-ended visual storytelling via latent diffusion models, 2024
Chang Liu, Haoning Wu, Yujie Zhong, Xiaoyun Zhang, Yanfeng Wang, and Weidi Xie · 2024
Closest in time.
Llmr: Real-time prompting of interactive worlds using large language models, 2024
Fernanda De La Torre, Cathy Mengying Fang, Han Huang, Andrzej Banburski-Fahey, Judith Amores Fernandez, and Jaron Lanier · 2024
Closest in time.
A multi-modal story generation framework with ai-driven storyline guidance
Juntae Kim, Yoonseok Heo, Hogeon Yu, and Jongho Nang · 2079
Closest in time.