Fetching the paper…
Reading the bibliography…
We present DreamPose, a diffusion-based method for generating animated fashion videos from still images.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, A.C. Bovik, H.R. Sheikh, and E.P. Simoncelli · 2004
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics, 2015
Jascha Sohl-Dickstein, Eric A. Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution, 2016
Justin Johnson, Alexandre Alahi, and Li Fei-Fei · 2016
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Earlier work this paper cites.
Densepose: Dense human pose estimation in the wild, 2018
Rıza Alp Güler, Natalia Neverova, and Iasonas Kokkinos · 2018
Earlier work this paper cites.
Animating arbitrary objects via deep motion transfer, 2018
Aliaksandr Siarohin, Stéphane Lathuilière, Sergey Tulyakov, Elisa Ricci, and Nicu Sebe · 2018
Earlier work this paper cites.
Towards accurate generative models of video: A new metric & challenges, 2018
Thomas Unterthiner, Sjoerd van Steenkiste, Karol Kurach, Raphael Marinier, Marcin Michalski, and Sylvain Gelly · 2018
Earlier work this paper cites.
Photo wake-up: 3d character animation from a single photo, 2018
Chung-Yi Weng, Brian Curless, and Ira Kemelmacher-Shlizerman · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric, 2018
Richard Zhang, Phillip Isola, Alexei A. Efros, Eli Shechtman, and Oliver Wang · 2018
Earlier work this paper cites.
Depth-aware video frame interpolation, 2019
Wenbo Bao, Wei-Sheng Lai, Chao Ma, Xiaoyun Zhang, Zhiyong Gao, and Ming-Hsuan Yang · 2019
Earlier work this paper cites.
Deepfashion2: A versatile benchmark for detection, pose estimation, segmentation and re-identification of clothing images, 2019
Yuying Ge, Ruimao Zhang, Lingyun Wu, Xiaogang Wang, Xiaoou Tang, and Ping Luo · 2019
Earlier work this paper cites.
Dwnet: Dense warp-based network for pose-guided human video generation, 2019
Polina Zablotskaia, Aliaksandr Siarohin, Bo Zhao, and Leonid Sigal · 2019
Earlier work this paper cites.
Progressive pose attention transfer for person image generation, 2019
Zhen Zhu, Tengteng Huang, Baoguang Shi, Miao Yu, Bofei Wang, and Xiang Bai · 2019
Earlier work this paper cites.
Animating pictures with eulerian motion fields, 2020
Aleksander Holynski, Brian Curless, Steven M. Seitz, and Richard Szeliski · 2020
Earlier work this paper cites.
Deep image spatial transformation for person image generation, 2020
Yurui Ren, Xiaoming Yu, Junming Chen, Thomas H. Li, and Ge Li · 2020
Earlier work this paper cites.
First order motion model for image animation, 2020
Aliaksandr Siarohin, Stéphane Lathuilière, Sergey Tulyakov, Elisa Ricci, and Nicu Sebe · 2020
Earlier work this paper cites.
Pose with style: Detail-preserving pose-guided image synthesis with conditional stylegan, 2021
Badour AlBahar, Jingwan Lu, Jimei Yang, Zhixin Shu, Eli Shechtman, and Jia-Bin Huang · 2021
Earlier work this paper cites.
Viton-hd: High-resolution virtual try-on via misalignment-aware normalization, 2021
Seunghwan Choi, Sunghyun Park, Minsoo Lee, and Jaegul Choo · 2021
Earlier work this paper cites.
Dressing in order: Recurrent person image generation for pose transfer, virtual try-on and outfit editing, 2021
Aiyu Cui, Daniel McKee, and Svetlana Lazebnik · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis, 2021
Prafulla Dhariwal and Alex Nichol · 2021
Cited alongside, same era.
Tryongan: Body-aware try-on via layered interpolation, 2021
Kathleen M Lewis, Srivatsan Varadharajan, and Ira Kemelmacher-Shlizerman · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Cited alongside, same era.
Motion representations for articulated animation
Aliaksandr Siarohin, Oliver J. Woodford, Jian Ren, Menglei Chai, and Sergey Tulyakov · 2021
Cited alongside, same era.
Pise: Person image synthesis and editing with decoupled gan, 2021
Jinsong Zhang, Kun Li, Yu-Kun Lai, and Jingyu Yang · 2021
Cited alongside, same era.
Person image synthesis via denoising diffusion model, 2022
Hierarchical text-conditional image generation with clip latents, 2022
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Later among the works it cites.
Dalle-2 is seeing double: Flaws in word-to-concept mapping in text2image models, 2022
Royi Rassin, Shauli Ravfogel, and Yoav Goldberg · 2022
Later among the works it cites.
Film: Frame interpolation for large motion, 2022
Fitsum Reda, Janne Kontkanen, Eric Tabellion, Deqing Sun, Caroline Pantofaru, and Brian Curless · 2022
Later among the works it cites.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation, 2022
Nataniel Ruiz, Yuanzhen Li, Varun Jampani, Yael Pritch, Michael Rubinstein, and Kfir Aberman · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding, 2022
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S. Sara Mahdavi, Rapha Gontijo Lopes, Tim Salimans, Jonathan Ho, David J Fleet, and Mohammad Norouzi · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ankan Kumar Bhunia, Salman Khan, Hisham Cholakkal, Rao Muhammad Anwer, Jorma Laaksonen, Mubarak Shah, and Fahad Shahbaz Khan · 2022
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions, 2022
Tim Brooks, Aleksander Holynski, and Alexei A. Efros · 2022
Cited alongside, same era.
Flexible diffusion modeling of long videos, 2022
William Harvey, Saeid Naderiparizi, Vaden Masrani, Christian Weilbach, and Frank Wood · 2022
Cited alongside, same era.
Latent video diffusion models for high-fidelity video generation with arbitrary lengths, 2022
Yingqing He, Tianyu Yang, Yong Zhang, Ying Shan, and Qifeng Chen · 2022
Cited alongside, same era.
Imagen video: High definition video generation with diffusion models, 2022
Jonathan Ho, William Chan, Chitwan Saharia, Jay Whang, Ruiqi Gao, Alexey Gritsenko, Diederik P. Kingma, Ben Poole, Mohammad Norouzi, David J. Fleet, and Tim Salimans · 2022
Cited alongside, same era.
Imagen video: High definition video generation with diffusion models, 2022
Jonathan Ho, William Chan, Chitwan Saharia, Jay Whang, Ruiqi Gao, Alexey Gritsenko, Diederik P. Kingma, Ben Poole, Mohammad Norouzi, David J. Fleet, and Tim Salimans · 2022
Cited alongside, same era.
Classifier-free diffusion guidance, 2022
Jonathan Ho and Tim Salimans · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding, 2022
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S. Sara Mahdavi, Rapha Gontijo Lopes, Tim Salimans, Jonathan Ho, David J Fleet, and Mohammad Norouzi · 2022
Later among the works it cites.
Make-a-video: Text-to-video generation without text-video data, 2022
Uriel Singer, Adam Polyak, Thomas Hayes, Xi Yin, Jie An, Songyang Zhang, Qiyuan Hu, Harry Yang, Oron Ashual, Oran Gafni, Devi Parikh, Sonal Gupta, and Yaniv Taigman · 2022
Later among the works it cites.
Latent image animator: Learning to animate images via latent space navigation, 2022
Yaohui Wang, Di Yang, Francois Bremond, and Antitza Dantcheva · 2022
Later among the works it cites.
Novel view synthesis with diffusion models, 2022
Daniel Watson, William Chan, Ricardo Martin-Brualla, Jonathan Ho, Andrea Tagliasacchi, and Mohammad Norouzi · 2022
Later among the works it cites.
Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation, 2022
Jay Zhangjie Wu, Yixiao Ge, Xintao Wang, Weixian Lei, Yuchao Gu, Wynne Hsu, Ying Shan, Xiaohu Qie, and Mike Zheng Shou · 2022
Later among the works it cites.
Diffusion probabilistic modeling for video generation, 2022
Ruihan Yang, Prakhar Srivastava, and Stephan Mandt · 2022
Later among the works it cites.
Thin-plate spline motion model for image animation, 2022
Jian Zhao and Hui Zhang · 2022
Later among the works it cites.
Universal guidance for diffusion models, 2023
Arpit Bansal, Hong-Min Chu, Avi Schwarzschild, Soumyadip Sengupta, Micah Goldblum, Jonas Geiping, and Tom Goldstein · 2023
Closest in time.
Difffashion: Reference-based fashion design with structure-aware transfer by diffusion models, 2023
Shidong Cao, Wenhao Chai, Shengyu Hao, Yanting Zhang, Hangyue Chen, and Gaoang Wang · 2023
Closest in time.
Designing an encoder for fast personalization of text-to-image models, 2023
Rinon Gal, Moab Arar, Yuval Atzmon, Amit H. Bermano, Gal Chechik, and Daniel Cohen-Or · 2023
Closest in time.
Dreamix: Video diffusion models are general video editors, 2023
Eyal Molad, Eliahu Horwitz, Dani Valevski, Alex Rav Acha, Yossi Matias, Yael Pritch, Yaniv Leviathan, and Yedid Hoshen · 2023
Closest in time.
Text-to-4d dynamic scene generation, 2023
Uriel Singer, Shelly Sheynin, Adam Polyak, Oron Ashual, Iurii Makarov, Filippos Kokkinos, Naman Goyal, Andrea Vedaldi, Devi Parikh, Justin Johnson, and Yaniv Taigman · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala · 2023
Closest in time.