Fetching the paper…
Reading the bibliography…
Image inpainting task refers to erasing unwanted pixels from images and filling them in a semantically consistent and realistic way.
Patchmatch: A randomized correspondence algorithm for structural image editing
C. Barnes, E. Shechtman, A. Finkelstein, and D. B. Goldman · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros · 2016
Earlier work this paper cites.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter · 2017
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. Van Der Maaten, L. Fei-Fei, C. Lawrence Zitnick, and R. Girshick · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Earlier work this paper cites.
Image inpainting for irregular holes using partial convolutions
G. Liu, F. A. Reda, K. J. Shih, T.-C. Wang, A. Tao, and B. Catanzaro · 2018
Earlier work this paper cites.
Cascade r-cnn: high quality object detection and instance segmentation
Z. Cai and N. Vasconcelos · 2019
Earlier work this paper cites.
Tell, draw, and repeat: Generating and modifying images based on continual linguistic instruction
A. El-Nouby, S. Sharma, H. Schulz, D. Hjelm, L. E. Asri, S. E. Kahou, Y. Bengio, and G. W. Taylor · 2019
Earlier work this paper cites.
LVIS: A dataset for large vocabulary instance segmentation
A. Gupta, P. Dollar, and R. Girshick · 2019
Earlier work this paper cites.
Gqa: A new dataset for real-world visual reasoning and compositional question answering
D. A. Hudson and C. D. Manning · 2019
Earlier work this paper cites.
CoDraw: Collaborative Drawing as a Testbed for Grounded Goal-driven Communication
J.-H. Kim, N. Kitaev, X. Chen, M. Rohrbach, Y. Tian, D. Batra, and D. Parikh · 2019
Earlier work this paper cites.
Panoptic feature pyramid networks
A. Kirillov, R. Girshick, K. He, and P. Dollár · 2019
Earlier work this paper cites.
Composing text and image for image retrieval-an empirical odyssey
N. Vo, L. Jiang, C. Sun, K. Murphy, L.-J. Li, L. Fei-Fei, and J. Hays · 2019
Earlier work this paper cites.
Detectron2
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick · 2019
Earlier work this paper cites.
Free-form image inpainting with gated convolution
J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang · 2019
Cited alongside, same era.
CascadePSP: Toward class-agnostic and very high-resolution segmentation via global and local refinement
H. K. Cheng, J. Chung, Y.-W. Tai, and C.-K. Tang · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
J. Ho, A. Jain, and P. Abbeel · 2020
Cited alongside, same era.
Recurrent feature reasoning for image inpainting
J. Li, N. Wang, L. Zhang, B. Du, and D. Tao · 2020
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo · 2021
Cited alongside, same era.
Sdedit: Guided image synthesis and editing with stochastic differential equations
C. Meng, Y. He, Y. Song, J. Song, J. Wu, J.-Y. Zhu, and S. Ermon · 2021
Mat: Mask-aware transformer for large hole image inpainting
W. Li, Z. Lin, K. Zhou, L. Qi, Y. Wang, and J. Jia · 2022
Later among the works it cites.
Partial convolution for padding, inpainting, and image synthesis
G. Liu, A. Dundar, K. J. Shih, T.-C. Wang, F. A. Reda, K. Sapra, Z. Yu, X. Yang, A. Tao, and B. Catanzaro · 2022
Later among the works it cites.
Image segmentation using text and image prompts
T. Lüddecke and A. Ecker · 2022
Later among the works it cites.
RePaint: Inpainting using denoising diffusion probabilistic models
A. Lugmayr, M. Danelljan, A. Romero, F. Yu, R. Timofte, and L. Van Gool · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
A. Ramesh, P. Dhariwal, A. Nichol, C. Chu, and M. Chen · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
A. Nichol, P. Dhariwal, A. Ramesh, P. Shyam, P. Mishkin, B. McGrew, I. Sutskever, and M. Chen · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever · 2021
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models, 2021
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2021
Cited alongside, same era.
Cr-fill: Generative image inpainting with auxiliary contextual reconstruction
Y. Zeng, Z. Lin, H. Lu, and V. M. Patel · 2021
Cited alongside, same era.
Text as neural operator: Image manipulation by text instruction
T. Zhang, H.-Y. Tseng, L. Jiang, W. Yang, H. Lee, and I. Essa · 2021
Cited alongside, same era.
Large scale image completion via co-modulated generative adversarial networks
S. Zhao, J. Cui, Y. Sheng, Y. Dong, X. Liang, E. I. Chang, and Y. Xu · 2021
Cited alongside, same era.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. Denton, S. K. S. Ghasemipour, B. K. Ayan, S. S. Mahdavi, R. G. Lopes, et al · 2022
Later among the works it cites.
Make-a-video: Text-to-video generation without text-video data
U. Singer, A. Polyak, T. Hayes, X. Yin, J. An, S. Zhang, Q. Hu, H. Yang, O. Ashual, O. Gafni, et al · 2022
Later among the works it cites.
High-fidelity guided image synthesis with latent diffusion models
J. Singh, S. Gould, and L. Zheng · 2022
Later among the works it cites.
Objectstitch: Generative object compositing
Y. Song, Z. Zhang, Z. Lin, S. Cohen, B. Price, J. Zhang, S. Y. Kim, and D. Aliaga · 2022
Later among the works it cites.
Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
J. Z. Wu, Y. Ge, X. Wang, W. Lei, Y. Gu, W. Hsu, Y. Shan, X. Qie, and M. Z. Shou · 2022
Later among the works it cites.
Scaling autoregressive models for content-rich text-to-image generation
J. Yu, Y. Xu, J. Y. Koh, T. Luong, G. Baid, Z. Wang, V. Vasudevan, A. Ku, Y. Yang, B. K. Ayan, B. Hutchinson, W. Han, Z. Parekh, X. Li, H. Zhang, J. Baldridge, and Y. Wu · 2022
Later among the works it cites.
High-fidelity image inpainting with gan inversion
Y. Yu, L. Zhang, H. Fan, and T. Luo · 2022
Later among the works it cites.
Detecting twenty-thousand classes using image-level supervision
X. Zhou, R. Girdhar, A. Joulin, P. Krähenbühl, and I. Misra · 2022
Later among the works it cites.
Generalized decoding for pixel, image, and language
X. Zou, Z.-Y. Dou, J. Yang, Z. Gan, L. Li, C. Li, X. Dai, H. Behl, J. Wang, L. Yuan, et al · 2022
Later among the works it cites.
Target-free text-guided image manipulation
W.-C. Fan, C.-F. Yang, C.-A. Yang, and Y.-C. F. Wang · 2023
Closest in time.
Diverse inpainting and editing with gan inversion
A. B. Yildirim, H. Pehlivan, B. B. Bilecen, and A. Dundar · 2023
Closest in time.