Fetching the paper…
Reading the bibliography…
Text-driven video editing has recently experienced rapid development.
Plug-and-play diffusion features for text-driven image-to-image translation
Tumanyan, N.; Geyer, M.; Bagon, S.; and Dekel, T. 2023 · 1930
Earlier work this paper cites.
Methodology for the Subjective Assessment of the Quality of Television Pictures ITU-R Recommendation
Int.Telecommun.Union. 2000 · 2000
Earlier work this paper cites.
Methodology for the subjective assessment of the quality of television pictures
Series, B. 2002 · 2002
Earlier work this paper cites.
A naturalistic open source movie for optical flow evaluation
Butler, D. J.; Wulff, J.; Stanley, G. B.; and Black, M. J. 2012 · 2012
Earlier work this paper cites.
AVA: A large-scale database for aesthetic visual analysis
Murray, N.; Marchesotti, L.; and Perronnin, F. 2012 · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2014 · 2014
Earlier work this paper cites.
The kinetics human action video dataset
Kay, W.; Carreira, J.; Simonyan, K.; Zhang, B.; Hillier, C.; Vijayanarasimhan, S.; Viola, F.; Green, T.; Back, T.; Natsev, P.; et al. 2017 · 2017
Earlier work this paper cites.
The 2017 davis challenge on video object segmentation
Pont-Tuset, J.; Perazzi, F.; Caelles, S.; Arbeláez, P.; Sorkine-Hornung, A.; and Van Gool, L. 2017 · 2017
Earlier work this paper cites.
Large-scale study of perceptual video quality
Sinno, Z.; and Bovik, A. C. 2018 · 2018
Earlier work this paper cites.
Towards accurate generative models of video: A new metric & challenges
Unterthiner, T.; Van Steenkiste, S.; Kurach, K.; Marinier, R.; Michalski, M.; and Gelly, S. 2018 · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Zhang, R.; Isola, P.; Efros, A. A.; Shechtman, E.; and Wang, O. 2018 · 2018
Earlier work this paper cites.
Learning to Rank for Blind Image Quality Assessment
Gao, F.; Tao, D.; Gao, X.; and Li, X. 2019 · 2019
Earlier work this paper cites.
Frozen in time: A joint video and image encoder for end-to-end retrieval
Bain, M.; Nagrani, A.; Varol, G.; and Zisserman, A. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Earlier work this paper cites.
Patch-vq:’patching up’the video quality problem
Ying, Z.; Mandal, M.; Ghadiyaram, D.; and Bovik, A. 2021 · 2021
Earlier work this paper cites.
Visual knowledge graph for human action reasoning in videos
Ma, Y.; Wang, Y.; Wu, Y.; Lyu, Z.; Chen, S.; Li, X.; and Qiao, Y. 2022 · 2022
Earlier work this paper cites.
Adavit: Adaptive vision transformers for efficient image recognition
Meng, L.; Li, H.; Chen, B.-C.; Lan, S.; Wu, Z.; Jiang, Y.-G.; and Lim, S.-N. 2022 · 2022
Earlier work this paper cites.
A deep learning based no-reference quality assessment model for ugc videos
Sun, W.; Min, X.; Lu, W.; and Zhai, G. 2022 · 2022
Earlier work this paper cites.
Fast-vqa: Efficient end-to-end video quality assessment with fragment sampling
Wu, H.; Chen, C.; Hou, J.; Liao, L.; Wang, A.; Sun, W.; Yan, Q.; and Lin, W. 2022 · 2022
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Brooks, T.; Holynski, A.; and Efros, A. A. 2023 · 2023
Cited alongside, same era.
Pix2video: Video editing using image diffusion
Ceylan, D.; Huang, C.-H. P.; and Mitra, N. J. 2023 · 2023
Cited alongside, same era.
Stablevideo: Text-driven consistency-aware diffusion video editing
Chai, W.; Guo, X.; Wang, G.; and Lu, Y. 2023 · 2023
Cited alongside, same era.
Text2video-zero: Text-to-image diffusion models are zero-shot video generators
Khachatryan, L.; Movsisyan, A.; Tadevosyan, V.; Henschel, R.; Wang, Z.; Navasardyan, S.; and Shi, H. 2023 · 2023
Cited alongside, same era.
Pick-a-pic: An open dataset of user preferences for text-to-image generation
Slicedit: Zero-Shot Video Editing With Text-to-Image Diffusion Models Using Spatio-Temporal Slices
Cohen, N.; Kulikov, V.; Kleiner, M.; Huberman-Spiegelglas, I.; and Michaeli, T. 2024 · 2024
Closest in time.
FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing
Cong, Y.; Xu, M.; Chen, S.; Ren, J.; Xie, Y.; Perez-Rua, J.-M.; Rosenhahn, B.; Xiang, T.; He, S.; et al. 2024 · 2024
Closest in time.
TokenFlow: Consistent Diffusion Features for Consistent Video Editing
Geyer, M.; Bar-Tal, O.; Bagon, S.; and Dekel, T. 2024 · 2024
Closest in time.
Evaluating Text to Image Synthesis: Survey and Taxonomy of Image Quality Metrics
Hartwig, S.; Engel, D.; Sick, L.; Kniesel, H.; Payer, T.; Ropinski, T.; et al. 2024 · 2024
Closest in time.
Diffusion model-based image editing: A survey
Huang, Y.; Huang, J.; Liu, Y.; Yan, M.; Lv, J.; Liu, J.; Xiong, W.; Zhang, H.; Chen, S.; and Cao, L. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kirstain, Y.; Polyak, A.; Singer, U.; Matiana, S.; Penna, J.; and Levy, O. 2023 · 2023
Cited alongside, same era.
Stablevqa: A deep no-reference quality assessment model for video stability
Kou, T.; Liu, X.; Sun, W.; Jia, J.; Min, X.; Zhai, G.; and Liu, N. 2023 · 2023
Cited alongside, same era.
Uniformer: Unifying convolution and self-attention for visual recognition
Li, K.; Wang, Y.; Zhang, J.; Gao, P.; Song, G.; Liu, Y.; Li, H.; and Qiao, Y. 2023 · 2023
Cited alongside, same era.
FETV: A Benchmark for Fine-Grained Evaluation of Open-Domain Text-to-Video Generation
Liu, Y.; Li, L.; Ren, S.; Gao, R.; Li, S.; Chen, S.; Sun, X.; and Hou, L. 2023 · 2023
Cited alongside, same era.
MagicStick: Controllable Video Editing via Control Handle Transformations
Ma, Y.; Cun, X.; He, Y.; Qi, C.; Wang, X.; Shan, Y.; Li, X.; and Chen, Q. 2023 · 2023
Cited alongside, same era.
Spring: A high-resolution high-detail dataset and benchmark for scene flow, optical flow and stereo
Mehl, L.; Schmalfuss, J.; Jahedi, A.; Nalivayko, Y.; and Bruhn, A. 2023 · 2023
Cited alongside, same era.
Fatezero: Fusing attentions for zero-shot text-based video editing
Qi, C.; Cun, X.; Zhang, Y.; Lei, C.; Wang, X.; Shan, Y.; and Chen, Q. 2023 · 2023
Cited alongside, same era.
Closest in time.
Rave: Randomized noise shuffling for fast and consistent video editing with diffusion models
Kara, O.; Kurtkaya, B.; Yesiltepe, H.; Rehg, J. M.; and Yanardag, P. 2024 · 2024
Closest in time.
Subjective-Aligned Dateset and Metric for Text-to-Video Quality Assessment
Kou, T.; Liu, X.; Zhang, Z.; Li, C.; Wu, H.; Min, X.; Zhai, G.; and Liu, N. 2024 · 2024
Closest in time.
Video-p2p: Video editing with cross-attention control
Liu, S.; Zhang, Y.; Li, W.; Lin, Z.; and Jia, J. 2024 · 2024
Closest in time.
DINOv2: Learning Robust Visual Features without Supervision
Oquab, M.; Darcet, T.; Moutakanni, T.; Vo, H. V.; Szafraniec, M.; Khalidov, V.; Fernandez, P.; HAZIZA, D.; Massa, F.; El-Nouby, A.; et al. 2024 · 2024
Closest in time.
Qu, B.; Liang, X.; Sun, S.; and Gao, W. 2024 · 2024
Closest in time.
RB-Modulation: Training-Free Personalization of Diffusion Models using Stochastic Optimal Control
Rout, L.; Chen, Y.; Ruiz, N.; Kumar, A.; Caramanis, C.; Shakkottai, S.; and Chu, W. 2024 · 2024
Closest in time.
T2v-compbench: A comprehensive benchmark for compositional text-to-video generation
Sun, K.; Huang, K.; Liu, X.; Wu, Y.; Xu, Z.; Li, Z.; and Liu, X. 2024 · 2024
Closest in time.
Chains of Diffusion Models
Wei, Y.; Huang, L.; Wu, Z.-F.; Wang, W.; Liu, Y.; Jia, M.; and Ma, S. 2024 · 2024
Closest in time.
Imagereward: Learning and evaluating human preferences for text-to-image generation
Xu, J.; Liu, X.; Wu, Y.; Tong, Y.; Li, Q.; Ding, M.; Tang, J.; and Dong, Y. 2024 · 2024
Closest in time.
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
Yang, S.; Zhou, Y.; Liu, Z.; and Loy, C. C. 2024 · 2024
Closest in time.
ControlVideo: Training-free Controllable Text-to-video Generation
Zhang, Y.; Wei, Y.; Jiang, D.; ZHANG, X.; Zuo, W.; and Tian, Q. 2024 · 2024
Closest in time.
InstantSwap: Fast Customized Concept Swapping across Sharp Shape Differences
Zhu, C.; Li, K.; Ma, Y.; Tang, L.; Fang, C.; Chen, C.; Chen, Q.; and Li, X. 2024 · 2024
Closest in time.