Fetching the paper…
Reading the bibliography…
Despite recent advances in UNet-based image editing, methods for shape-aware object editing in high-resolution images are still lacking.
Multi-concept customization of text-to-image diffusion
Kumari, N.; Zhang, B.; Zhang, R.; Shechtman, E.; and Zhu, J.-Y. 2023 · 1941
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; et al. 2020 · 2010
Earlier work this paper cites.
Denoising Diffusion Implicit Models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2010
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Earlier work this paper cites.
Photographic text-to-image synthesis with a hierarchically-nested adversarial network
Zhang, Z.; Xie, Y.; and Yang, L. 2018 · 2018
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Cpgan: Content-parsing generative adversarial networks for text-to-image synthesis
Liang, J.; Pei, W.; and Lu, F. 2020 · 2020
Earlier work this paper cites.
Ilvr: Conditioning method for denoising diffusion probabilistic models
Choi, J.; Kim, S.; Jeong, Y.; Gwon, Y.; and Yoon, S. 2021 · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal, P.; and Nichol, A. 2021 · 2021
Earlier work this paper cites.
Taming transformers for high-resolution image synthesis
Esser, P.; Rombach, R.; and Ommer, B. 2021 · 2021
Earlier work this paper cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Meng, C.; He, Y.; Song, Y.; Song, J.; Wu, J.; Zhu, J.-Y.; and Ermon, S. 2021 · 2021
Earlier work this paper cites.
Denoising Diffusion Implicit Models
Song, J.; Meng, C.; and Ermon, S. 2021 · 2021
Earlier work this paper cites.
The stable artist: Steering semantics in diffusion latent space
Brack, M.; Schramowski, P.; Friedrich, F.; Hintersdorf, D.; and Kersting, K. 2022 · 2022
Earlier work this paper cites.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Gal, R.; Alaluf, Y.; Atzmon, Y.; Patashnik, O.; Bermano, A. H.; Chechik, G.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Prompt-to-prompt image editing with cross attention control
Hertz, A.; Mokady, R.; Tenenbaum, J.; Aberman, K.; Pritch, Y.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Lu, C.; Zhou, Y.; Bao, F.; Chen, J.; Li, C.; and Zhu, J. 2022 · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Cited alongside, same era.
Fast sampling of diffusion models with exponential integrator
MagicStick: Controllable Video Editing via Control Handle Transformations
Ma, Y.; Cun, X.; He, Y.; Qi, C.; Wang, X.; Shan, Y.; Li, X.; and Chen, Q. 2023 · 2023
Later among the works it cites.
NULL-Text Inversion for Editing Real Images Using Guided Diffusion Models
Mokady, R.; Hertz, A.; Aberman, K.; Pritch, Y.; and Cohen-Or, D. 2023 · 2023
Later among the works it cites.
Zero-shot image-to-image translation
Parmar, G.; Kumar Singh, K.; Zhang, R.; Li, Y.; Lu, J.; and Zhu, J.-Y. 2023 · 2023
Later among the works it cites.
Scalable Diffusion Models with Transformers
Peebles, W.; and Xie, S. 2023 · 2023
Later among the works it cites.
Inversion-free image editing with natural language
Xu, S.; Huang, Y.; Pan, J.; Ma, Z.; and Chai, J. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhang, Q.; and Chen, Y. 2022 · 2022
Cited alongside, same era.
Improving image generation with better captions
Betker, J.; Goh, G.; Jing, L.; Brooks, T.; Wang, J.; Li, L.; Ouyang, L.; Zhuang, J.; Lee, J.; Guo, Y.; et al. 2023 · 2023
Cited alongside, same era.
Token Merging: Your ViT But Faster
Bolya, D.; Fu, C.-Y.; Dai, X.; Zhang, P.; Feichtenhofer, C.; and Hoffman, J. 2023 · 2023
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Brooks, T.; Holynski, A.; and Efros, A. A. 2023 · 2023
Cited alongside, same era.
Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing
Cao, M.; Wang, X.; Qi, Z.; Shan, Y.; Qie, X.; and Zheng, Y. 2023 · 2023
Cited alongside, same era.
PixArt- α \alpha : Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
Chen, J.; Yu, J.; Ge, C.; Yao, L.; Xie, E.; Wu, Y.; Wang, Z.; Kwok, J.; Luo, P.; Lu, H.; et al. 2023 · 2023
Cited alongside, same era.
Prompt Tuning Inversion for Text-driven Image Editing Using Diffusion Models
Dong, W.; Xue, S.; Duan, X.; and Han, S. 2023 · 2023
Cited alongside, same era.
Scalecrafter: Tuning-free higher-resolution visual generation with diffusion models
He, Y.; Yang, S.; Chen, H.; Cun, X.; Xia, M.; Zhang, Y.; Wang, X.; He, R.; Chen, Q.; and Shan, Y. 2023 · 2023
Cited alongside, same era.
Adding conditional control to text-to-image diffusion models
Zhang, L.; Rao, A.; and Agrawala, M. 2023 · 2023
Later among the works it cites.
SINE: SINgle Image Editing With Text-to-Image Diffusion Models
Zhang, Z.; Han, L.; Ghosh, A.; Metaxas, D. N.; and Ren, J. 2023 · 2023
Later among the works it cites.
Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation
Chen, Q.; Ma, Y.; Wang, H.; Yuan, J.; Zhao, W.; Tian, Q.; Wang, H.; Min, S.; Chen, Q.; and Liu, W. 2024 · 2024
Closest in time.
On Exact Inversion of DPM-Solvers
Hong, S.; Lee, K.; Jeon, S. Y.; Bae, H.; and Chun, S. Y. 2024 · 2024
Closest in time.
Pnp inversion: Boosting diffusion-based editing with 3 lines of code
Ju, X.; Zeng, A.; Bian, Y.; Liu, S.; and Xu, Q. 2024 · 2024
Closest in time.
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
Wang, J.; Ma, Y.; Guo, J.; Xiao, Y.; Huang, G.; and Li, X. 2024 · 2024
Closest in time.
Inversion-Free Image Editing with Language-Guided Diffusion Models
Xu, S.; Huang, Y.; Pan, J.; Ma, Z.; and Chai, J. 2024 · 2024
Closest in time.
Improving diffusion-based image synthesis with context prediction
Yang, L.; Liu, J.; Hong, S.; Zhang, Z.; Huang, Z.; Cai, Z.; Zhang, W.; and Cui, B. 2024 · 2024
Closest in time.