Fetching the paper…
Reading the bibliography…
The growing demand for high-fidelity video generation from textual descriptions has catalyzed significant research in this field.
Multiple video frame interpolation via enhanced deformable separable convolution, 2021
Xianhang Cheng and Zhenzhong Chen · 2021
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Earlier work this paper cites.
https://research.runwayml.com/gen2
Gen-2 · 2023
Earlier work this paper cites.
https://https://moonvalley.ai/
MoonValley · 2023
Earlier work this paper cites.
https://www.morphstudio.com/
Morph · 2023
Earlier work this paper cites.
https://pika.art/
Pika 1.0 · 2023
Cited alongside, same era.
https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt
SVD-XT · 2023
Cited alongside, same era.
Stable video diffusion: Scaling latent video diffusion models to large datasets, 2023
Andreas Blattmann, Tim Dockhorn, Sumith Kulal, Daniel Mendelevitch, Maciej Kilian, Dominik Lorenz, Yam Levi, Zion English, Vikram Voleti, Adam Letts, Varun Jampani, and Robin Rombach · 2023
Cited alongside, same era.
Ldmvfi: Video frame interpolation with latent diffusion models, 2023
Duolikun Danier, Fan Zhang, and David Bull · 2023
Cited alongside, same era.
Extracting motion and appearance via inter-frame attention for efficient video frame interpolation
Guozhen Zhang, Yuhan Zhu, Haonan Wang, Youxin Chen, Gangshan Wu, and Limin Wang
Cited in the paper.
Adding conditional control to text-to-image diffusion models, 2023b
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala
Cited in the paper.
Emu video: Factorizing text-to-video generation by explicit image conditioning, 2023
Rohit Girdhar, Mannat Singh, Andrew Brown, Quentin Duval, Samaneh Azadi, Sai Saketh Rambhatla, Akbar Shah, Xi Yin, Devi Parikh, and Ishan Misra · 2023
Later among the works it cites.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning, 2023
Yuwei Guo, Ceyuan Yang, Anyi Rao, Yaohui Wang, Yu Qiao, Dahua Lin, and Bo Dai · 2023
Later among the works it cites.
Videopoet: A large language model for zero-shot video generation, 2023
Dan Kondratyuk, Lijun Yu, Xiuye Gu, José Lezama, Jonathan Huang, Rachel Hornung, Hartwig Adam, Hassan Akbari, Yair Alon, Vighnesh Birodkar, Yong Cheng, Ming-Chang Chiu, Josh Dillon, Irfan Essa, Agrim Gupta, Meera Hahn, Anja Hauth, David Hendon, Alonso Martinez, David Minnen, David Ross, Grant Schindler, Mikhail Sirotenko, Kihyuk Sohn, Krishna Somandepalli, Huisheng Wang, Jimmy Yan, Ming-Hsuan Yang, Xuan Yang, Bryan Seybold, and Lu Jiang · 2023
Later among the works it cites.
Magicvideo: Efficient video generation with latent diffusion models, 2023
Daquan Zhou, Weimin Wang, Hanshu Yan, Weiwei Lv, Yizhe Zhu, and Jiashi Feng · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…