Fetching the paper…
Reading the bibliography…
Image editing serves as a practical yet challenging task considering the diverse demands from users, where one of the hardest parts is to precisely describe how the edited image should look like.
Object recognition from local scale-invariant features
D. G. Lowe · 1999
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli · 2004
Earlier work this paper cites.
Image quality metrics: Psnr vs. ssim
A. Hore and D. Ziou · 2010
Earlier work this paper cites.
Multi-scale image harmonization
K. Sunkavalli, M. K. Johnson, W. Matusik, and H. Pfister · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Image inpainting for irregular holes using partial convolutions
G. Liu, F. A. Reda, K. J. Shih, T.-C. Wang, A. Tao, and B. Catanzaro · 2018
Earlier work this paper cites.
Generative image inpainting with contextual attention
J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang · 2018
Earlier work this paper cites.
Toward realistic image compositing with adversarial learning
B.-C. Chen and A. Kae · 2019
Earlier work this paper cites.
Free-form image inpainting with gated convolution
J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang · 2019
Earlier work this paper cites.
Dovenet: Deep image harmonization via domain verification
W. Cong, J. Zhang, L. Niu, L. Liu, Z. Ling, W. Li, and L. Zhang · 2020
Earlier work this paper cites.
Intrinsic image harmonization
Z. Guo, H. Zheng, Y. Jiang, Z. Gu, and B. Zheng · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Hierarchical dynamic image harmonization
H. Chen, Z. Gu, Y. Li, J. Lan, C. Meng, W. Wang, and H. Li · 2022
Earlier work this paper cites.
High-resolution image harmonization via collaborative dual transformations
W. Cong, X. Tao, L. Niu, J. Liang, X. Gao, Q. Sun, and L. Zhang · 2022
Earlier work this paper cites.
Mat: Mask-aware transformer for large hole image inpainting
W. Li, Z. Lin, K. Zhou, L. Qi, Y. Wang, and J. Jia · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, et al · 2022
Cited alongside, same era.
Break-a-scene: Extracting multiple concepts from a single image
O. Avrahami, K. Aberman, O. Fried, D. Cohen-Or, and D. Lischinski · 2023
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
T. Brooks, A. Holynski, and A. A. Efros · 2023
Cited alongside, same era.
Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing
M. Cao, X. Wang, Z. Qi, Y. Shan, X. Qie, and Y. Zheng · 2023
Cited alongside, same era.
Livephoto: Real image animation with text-guided motion control
Emergent correspondence from image diffusion
L. Tang, M. Jia, Q. Wang, C. P. Phoo, and B. Hariharan · 2023
Later among the works it cites.
Paint by example: Exemplar-based image editing with diffusion models
B. Yang, S. Gu, B. Zhang, T. Zhang, X. Chen, X. Sun, D. Chen, and F. Wen · 2023
Later among the works it cites.
Ip-adapter: Text compatible image prompt adapter for text-to-image diffusion models
H. Ye, J. Zhang, S. Liu, X. Han, and W. Yang · 2023
Later among the works it cites.
Inpaint anything: Segment anything meets image inpainting
T. Yu, R. Feng, R. Feng, J. Liu, X. Jin, W. Zeng, and Z. Chen · 2023
Later among the works it cites.
Customnet: Zero-shot object customization with variable-viewpoints in text-to-image diffusion models
Z. Yuan, M. Cao, X. Wang, Z. Qi, C. Yuan, and Y. Shan · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X. Chen, Z. Liu, M. Chen, Y. Feng, Y. Liu, Y. Shen, and H. Zhao · 2023
Cited alongside, same era.
An image is worth one word: Personalizing text-to-image generation using textual inversion
R. Gal, Y. Alaluf, Y. Atzmon, O. Patashnik, A. H. Bermano, G. Chechik, and D. Cohen-Or · 2023
Cited alongside, same era.
Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models
Y. Gu, X. Wang, J. Z. Wu, Y. Shi, Y. Chen, Z. Fan, W. Xiao, R. Zhao, S. Chang, W. Wu, et al · 2023
Cited alongside, same era.
Prompt-to-prompt image editing with cross-attention control
A. Hertz, R. Mokady, J. Tenenbaum, K. Aberman, Y. Pritch, and D. Cohen-or · 2023
Cited alongside, same era.
Imagic: Text-based real image editing with diffusion models
B. Kawar, S. Zada, O. Lang, O. Tov, H. Chang, T. Dekel, I. Mosseri, and M. Irani · 2023
Cited alongside, same era.
Segment anything
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, et al · 2023
Cited alongside, same era.
Null-text inversion for editing real images using guided diffusion models
R. Mokady, A. Hertz, K. Aberman, Y. Pritch, and D. Cohen-Or · 2023
Cited alongside, same era.
Later among the works it cites.
Reference-only controlnet
L. Zhang · 2023
Later among the works it cites.
Swapanything: Enabling arbitrary object swapping in personalized image editing
J. Gu, Y. Wang, N. Zhao, W. Xiong, Q. Liu, Z. Zhang, H. Zhang, J. Zhang, H. Jung, and X. E. Wang · 2024
Closest in time.
Animate anyone: Consistent and controllable image-to-video synthesis for character animation
L. Hu, X. Gao, P. Zhang, K. Sun, B. Zhang, and L. Bo · 2024
Closest in time.
Dinov2: Learning robust visual features without supervision
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, et al · 2024
Closest in time.
Locate, assign, refine: Taming customized image inpainting with text-subject guidance
Y. Pan, C. Mao, Z. Jiang, Z. Han, and J. Zhang · 2024
Closest in time.
The best free stock photos, royalty free images & videos shared by creators
Pexels · 2024
Closest in time.
Clic: Concept learning in context
M. Safaee, A. Mikaeili, O. Patashnik, D. Cohen-Or, and A. Mahdavi-Amiri · 2024
Closest in time.
Imprint: Generative object compositing by learning identity-preserving representation
Y. Song, Z. Zhang, Z. Lin, S. Cohen, B. Price, J. Zhang, S. Y. Kim, H. Zhang, W. Xiong, and D. Aliaga · 2024
Closest in time.
Realfill: Reference-driven generation for authentic image completion
L. Tang, N. Ruiz, Q. Chu, Y. Li, A. Holynski, D. E. Jacobs, B. Hariharan, Y. Pritch, N. Wadhwa, K. Aberman, et al · 2024
Closest in time.
Depth anything: Unleashing the power of large-scale unlabeled data
L. Yang, B. Kang, Z. Huang, X. Xu, J. Feng, and H. Zhao · 2024
Closest in time.
Flashface: Human image personalization with high-fidelity identity preservation
S. Zhang, L. Huang, X. Chen, Y. Zhang, Z.-F. Wu, Y. Feng, W. Wang, Y. Shen, Y. Liu, and P. Luo · 2024
Closest in time.