Fetching the paper…
Reading the bibliography…
Recent advances in AI-driven storytelling have enhanced video generation and story visualization.
Reynal & Hitchcock, 1943
A. de Saint-Exupéry, The Little Prince · 1943
Earlier work this paper cites.
Burlington, Ma: Focal Press, 2010
K. Reisz, G. Millar, and E. Al, The technique of film editing · 2010
Earlier work this paper cites.
A. Mittal, R. Soundararajan, and A. C. Bovik, “Making a “completely blind” image quality analyzer,” IEEE Signal processing letters
2013
Earlier work this paper cites.
Y. Li, Z. Gan, Y. Shen, J. Liu, Y. Cheng, Y. Wu, L. Carin, D. Carlson, and J. Gao, “Storygan: A sequential conditional gan for story visualization,” in CVPR
2019
Earlier work this paper cites.
Y. Song, Z. R. Tam, H. Chen, H. Lu, and H. Shuai, “Character-preserving coherent story visualization,” in ECCV
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel, et al
2020
Earlier work this paper cites.
A. Maharana and M. Bansal, “Integrating visuospatial, linguistic, and commonsense structure into story visualization,” in EMNLP
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al
2021
Earlier work this paper cites.
B. Li and T. Lukasiewicz, “Learning to model multimodal semantic alignment for story visualization,” in EMNLP
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in CVPR
2022
Earlier work this paper cites.
X. Miao, G. Oliaro, Z. Zhang, X. Cheng, H. Jin, T. Chen, and Z. Jia, “Towards efficient generative large language model serving: A survey from algorithms to systems,” 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Y. Gong, Y. Pang, X. Cun, M. Xia, Y. He, H. Chen, L. Wang, Y. Zhang, X. Wang, Y. Shan, and Y. Yang, “Interactive story visualization with multiple characters,” in SIGGRAPH Asia 2023 Conference Papers
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
T. Brooks, B. Peebles, C. Holmes, W. DePue, Y. Guo, L. Jing, D. Schnurr, J. Taylor, T. Luhman, E. Luhman, C. Ng, R. Wang, and A. Ramesh, “Video generation models as world simulators,” 2024
2024
Cited alongside, same era.
Z. Xie, H. Mo, and C. Gao, “Video-driven sketch animation via cyclic reconstruction mechanism,” in ICME
2024
Cited alongside, same era.
Y. Guo, R. Yan, Y. Wu, and S. Ma, “Styleself: Style-controllable high-fidelity conversational virtual avatars generation,” in ICMEW
2024
Cited alongside, same era.
J. Guo, S. Su, J. Zhu, L. Gao, and J. Song, “Training-free semantic video composition via pre-trained diffusion model,” in ICME
2024
Cited alongside, same era.
F. Shen and J. Tang, “Imagpose: A unified conditional framework for pose-guided person generation,” in NeuralIPS
2024
Closest in time.
L. Yang, Z. Yu, C. Meng, M. Xu, S. Ermon, and B. Cui, “Mastering text-to-image diffusion: Recaptioning, planning, and generating with multimodal llms,” in ICML
2024
Closest in time.
2024
Closest in time.
W. Feng, W. Zhu, T.-j. Fu, V. Jampani, A. Akula, X. He, S. Basu, X. E. Wang, and W. Y. Wang, “Layoutgpt: Compositional visual planning and generation with large language models,” NeurIPS
2024
Closest in time.
Y. Li, H. Shi, B. Hu, L. Wang, J. Zhu, J. Xu, Z. Zhao, and M. Zhang, “Anim-director: A large multimodal model powered agent for controllable animation video generation,” in SIGGRAPH Asia 2024 Conference Papers
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
L. Zhang, A. Rao, and M. Agrawala, “Ic-light github page,” 2024
2024
Cited alongside, same era.
Q. Huang, S. Fu, J. Liu, H. Jiang, Y. Yu, and J. Song, “Resolving multi-condition confusion for finetuning-free personalized image generation,” 2024
2024
Cited alongside, same era.
Z. Zhou, J. Li, H. Li, N. Chen, and X. Tang, “Storymaker: Towards holistic consistent characters in text-to-image generation,” 2024
2024
Cited alongside, same era.
Y. Zhou, D. Zhou, M.-M. Cheng, J. Feng, and Q. Hou, “Storydiffusion: Consistent self-attention for long-range image and video generation,” in NeuralIPS
2024
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. H. Chi, Q. V. Le, and D. Zhou, “Chain-of-thought prompting elicits reasoning in large language models,” in NeuralIPS
2024
Cited alongside, same era.
2024
Closest in time.
X. Pan, P. Qin, Y. Li, H. Xue, and W. Chen, “Synthesizing coherent story with auto-regressive latent diffusion models,” in WACV
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
X. Yang, H. Shi, B. Zhang, F. Yang, J. Wang, H. Zhao, X. Liu, X. Wang, Q. Lin, J. Yu, L. Wang, Z. Chen, S. Liu, Y. Liu, Y. Yang, D. Wang, J. Jiang, and C. Guo, “Tencent hunyuan3d-1.0: A unified framework for text-to-3d and image-to-3d generation,” 2024
2024
Closest in time.
Accessed: 2024-12-21
OpenAI, “Chatgpt (version 4),” 2024 · 2024
Closest in time.
Accessed: 2024-12-22
Civitai, “Dream creation: Virtual 3d or e-commerce scene key visual poster or blind box ip display c4d super visual.” Civitai, 2024 · 2024
Closest in time.
Accessed: 2024-12-22
OpenAI, “Dall·e 3: Text-to-image generation.” OpenAI, 2024 · 2024
Closest in time.