Fetching the paper…
Reading the bibliography…
Humans watch more than a billion hours of video per day.
“Few-shot Video-to-Video Synthesis”, 2019
Ting-Chun Wang et al · 1910
Earlier work this paper cites.
“Collecting highly parallel data for paraphrase evaluation”
David Chen and William Dolan · 2011
Earlier work this paper cites.
“A Benchmark Dataset and Evaluation Methodology for Video Object Segmentation”
F. Perazzi et al · 2016
Earlier work this paper cites.
“The 2017 DAVIS Challenge on Video Object Segmentation”
Jordi Pont-Tuset et al · 2017
Earlier work this paper cites.
“Video-to-Video Synthesis”, 2018
Ting-Chun Wang et al · 2018
Earlier work this paper cites.
“FVD: A new metric for video generation”
Thomas Unterthiner et al · 2019
Earlier work this paper cites.
“CLIPScore: A Reference-free Evaluation Metric for Image Captioning”
Jack Hessel et al · 2021
Earlier work this paper cites.
“LoRA: Low-Rank Adaptation of Large Language Models”, 2021
Edward. Hu et al · 2021
Earlier work this paper cites.
“OpenCLIP” If you use this software, please cite it as below
Gabriel Ilharco et al · 2021
Earlier work this paper cites.
“DreamFusion: Text-to-3D using 2D Diffusion”
Ben Poole, Ajay Jain, Jonathan. Barron and Ben Mildenhall · 2022
Earlier work this paper cites.
“Imagen video: High definition video generation with diffusion models”
Jonathan Ho et al · 2022
Earlier work this paper cites.
“Make-A-Video: Text-to-Video Generation without Text-Video Data”, 2022
Uriel Singer et al · 2022
Earlier work this paper cites.
“Imagen video: High definition video generation with diffusion models”
Jonathan Ho et al · 2022
Earlier work this paper cites.
“Prompt-to-Prompt Image Editing with Cross Attention Control”, 2022
Amir Hertz et al · 2022
Earlier work this paper cites.
“Text2live: Text-driven layered image and video editing”
Omer Bar-Tal et al · 2022
Cited alongside, same era.
Jonathan Ho et al · 2022
Cited alongside, same era.
“Latent Video Diffusion Models for High-Fidelity Long Video Generation”, 2022
Yingqing He et al · 2022
Cited alongside, same era.
“Dreamix: Video Diffusion Models are General Video Editors”
Eyal Molad et al · 2023
Cited alongside, same era.
“Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation”
Jay Wu et al · 2023
Cited alongside, same era.
“TokenFlow: Consistent Diffusion Features for Consistent Video Editing”, 2023
Michal Geyer, Omer Bar-Tal, Shai Bagon and Tali Dekel · 2023
Closest in time.
“MotionDirector: Motion Customization of Text-to-Video Diffusion Models”
Rui Zhao et al · 2023
Closest in time.
Jia-Wei Liu et al · 2023
Closest in time.
“Audioldm: Text-to-audio generation with latent diffusion models”
Haohe Liu et al · 2023
Closest in time.
“Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation”, 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Patrick Esser et al · 2023
Cited alongside, same era.
“Edit-A-Video: Single Video Editing with Object-Aware Consistency”
Chaehun Shin et al · 2023
Cited alongside, same era.
“Video-P2P: Video Editing with Cross-attention Control”
Shaoteng Liu et al · 2023
Cited alongside, same era.
“Zero-shot video editing using off-the-shelf image diffusion models”
Wen Wang et al · 2023
Cited alongside, same era.
“FateZero: Fusing Attentions for Zero-shot Text-based Video Editing”, 2023
Chenyang Qi et al · 2023
Cited alongside, same era.
“Pix2Video: Video Editing using Image Diffusion”, 2023
Duygu Ceylan, Chun-Hao Huang and Niloy. Mitra · 2023
Cited alongside, same era.
“Text2Video-Zero: Text-to-Image Diffusion Models are Zero-Shot Video Generators”, 2023
Levon Khachatryan et al · 2023
Cited alongside, same era.
David Zhang et al · 2023
Closest in time.
“Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models”, 2023
Andreas Blattmann et al · 2023
Closest in time.
“InstructPix2Pix: Learning to Follow Image Editing Instructions”
Tim Brooks, Aleksander Holynski and Alexei. Efros · 2023
Closest in time.
“Make-A-Protagonist: Generic Video Editing with An Ensemble of Experts”
Yuyang Zhao et al · 2023
Closest in time.
“ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation”, 2023
Jiazheng Xu et al · 2023
Closest in time.
“Pick-a-pic: An open dataset of user preferences for text-to-image generation”
Yuval Kirstain et al · 2023
Closest in time.
“Adding Conditional Control to Text-to-Image Diffusion Models”, 2023
Lvmin Zhang and Maneesh Agrawala · 2023
Closest in time.
Alexander Kirillov et al · 2023
Closest in time.
“DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation”, 2023
Nataniel Ruiz et al · 2023
Closest in time.