Fetching the paper…
Reading the bibliography…
In an era where images and visual content dominate our digital landscape, the ability to manipulate and personalize these images has become a necessity.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T. (2015) · 2015
Earlier work this paper cites.
Visual atribute transfer through deep image analogy
Liao, J., Yao, Y., Yuan, L., Hua, G., and Kang, S. B. (2017) · 2017
Earlier work this paper cites.
Large scale gan training for high fidelity natural image synthesis
Brock, A., Donahue, J., and Simonyan, K. (2018) · 2018
Earlier work this paper cites.
Multimodal unsupervised image-to-image translation
Huang, X., Liu, M.-Y., Belongie, S., and Kautz, J. (2018) · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Karras, T., Laine, S., and Aila, T. (2019) · 2019
Earlier work this paper cites.
Example-guided style-consistent image synthesis from semantic labeling
Wang, M., Yang, G.-Y., Li, R., Liang, R.-Z., Zhang, S.-H., Hall, P. M., and Hu, S.-M. (2019) · 2019
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2020) · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S. (2020) · 2020
Earlier work this paper cites.
Cross-domain correspondence learning for exemplar-based image translation
Zhang, P., Zhang, B., Chen, D., Yuan, L., and Wen, F. (2020) · 2020
Earlier work this paper cites.
Instance-Conditioned GAN
Casanova, A., Careil, M., Verbeek, J., Drozdzal, M., and Romero-Soriano, A. (2021) · 2021
Earlier work this paper cites.
Cogview: Mastering text-to-image generation via transformers
Ding, M., Yang, Z., Hong, W., Zheng, W., Zhou, C., Yin, D., Lin, J., Zou, X., Shao, Z., Yang, H., · 2021
Earlier work this paper cites.
High-resolution complex scene synthesis with transformers
Jahn, M., Rombach, R., and Ommer, B. (2021) · 2021
Earlier work this paper cites.
Adaattn: Revisit attention mechanism in arbitrary neural style transfer
Liu, S., Lin, T., He, D., Li, F., Wang, M., Li, X., Sun, Z., Li, Q., and Ding, E. (2021) · 2021
Earlier work this paper cites.
Sdedit: Image synthesis and editing with stochastic differential equations
Meng, C., Song, Y., Song, J., Wu, J., Zhu, J.-Y., and Ermon, S. (2021) · 2021
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A., Dhariwal, P., Ramesh, A., Shyam, P., Mishkin, P., McGrew, B., Sutskever, I., and Chen, M. (2021) · 2021
Earlier work this paper cites.
DALL·E: Creating images from text
OpenAI (2021) · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., · 2021
Cited alongside, same era.
Cocosnet v2: Full-resolution correspondence learning for image translation
Zhou, X., Zhang, B., Zhang, T., Zhang, P., Bao, J., Chen, D., Zhang, Z., and Wen, F. (2021) · 2021
Cited alongside, same era.
Retrieval-Augmented Diffusion Models
Blattmann, A., Rombach, R., Oktay, K., Müller, J., and Ommer, B. (2022) · 2022
Cited alongside, same era.
Diffedit: Diffusion-based semantic image editing with mask guidance
Couairon, G., Verbeek, J., Schwenk, H., and Cord, M. (2022) · 2022
Cited alongside, same era.
Vqgan-clip: Open domain image generation and editing with natural language guidance
Crowson, K., Biderman, S., Kornis, D., Stander, D., Hallahan, E., Castricato, L., and Raff, E. (2022) · 2022
Midms: Matching interleaved diffusion models for exemplar-based image translation
Seo, J., Lee, G., Cho, S., Lee, J., and Kim, S. (2022) · 2022
Later among the works it cites.
Plug-and-play diffusion features for text-driven image-to-image translation
Tumanyan, N., Geyer, M., Bagon, S., and Dekel, T. (2022) · 2022
Later among the works it cites.
Scenecomposer: Any-level semantic image synthesis
Zeng, Y., Lin, Z., Zhang, J., Liu, Q., Collomosse, J., Kuen, J., and Patel, V. M. (2022) · 2022
Later among the works it cites.
Inversion-based creativity transfer with diffusion models
Zhang, Y., Huang, N., Tang, F., Huang, H., Ma, C., Dong, W., and Xu, C. (2022) · 2022
Later among the works it cites.
Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing
Cao, M., Wang, X., Qi, Z., Shan, Y., Qie, X., and Zheng, Y. (2023) · 2023
Closest in time.
Re-Imagen: Retrieval-Augmented Text-to-Image Generator
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Stytr2: Image style transfer with transformers
Deng, Y., Tang, F., Dong, W., Ma, C., Pan, X., Wang, L., and Xu, C. (2022) · 2022
Cited alongside, same era.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Gal, R., Alaluf, Y., Atzmon, Y., Patashnik, O., Bermano, A. H., Chechik, G., and Cohen-Or, D. (2022) · 2022
Cited alongside, same era.
Vector quantized diffusion model for text-to-image synthesis
Gu, S., Chen, D., Bao, J., Wen, F., Zhang, B., Chen, D., Yuan, L., and Guo, B. (2022) · 2022
Cited alongside, same era.
Prompt-to-prompt image editing with cross attention control
Hertz, A., Mokady, R., Tenenbaum, J., Aberman, K., Pritch, Y., and Cohen-Or, D. (2022) · 2022
Cited alongside, same era.
Imagic: Text-based real image editing with diffusion models
Kawar, B., Zada, S., Lang, O., Tov, O., Chang, H., Dekel, T., Mosseri, I., and Irani, M. (2022) · 2022
Cited alongside, same era.
Null-text Inversion for Editing Real Images using Guided Diffusion Models
Mokady, R., Hertz, A., Aberman, K., Pritch, Y., and Cohen-Or, D. (2022) · 2022
Cited alongside, same era.
Chen, W., Hu, H., Saharia, C., and Cohen, W. W. (2023) · 2023
Closest in time.
Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis
Feng, W., He, X., Fu, T.-J., Jampani, V., Akula, A., Narayana, P., Basu, S., Wang, X. E., and Wang, W. Y. (2023) · 2023
Closest in time.
Multi-Concept Customization of Text-to-Image Diffusion
Kumari, N., Zhang, B., Zhang, R., Shechtman, E., and Zhu, J.-Y. (2023) · 2023
Closest in time.
Gligen: Open-set grounded text-to-image generation
Li, Y., Liu, H., Wu, Q., Mu, F., Yang, J., Gao, J., Li, C., and Lee, Y. J. (2023) · 2023
Closest in time.
Analyzing bias in diffusion-based face generation models
Perera, M. V. and Patel, V. M. (2023) · 2023
Closest in time.
DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation
Ruiz, N., Li, Y., Jampani, V., Pritch, Y., Rubinstein, M., and Aberman, K. (2023) · 2023
Closest in time.
Stable bias: Analyzing societal representations in diffusion models
Sasha Luccioni, A., Akiki, C., Mitchell, M., and Jernite, Y. (2023) · 2023
Closest in time.
KNN-Diffusion: Image Generation via Large-Scale Retrieval
Sheynin, S., Ashual, O., Polyak, A., Singer, U., Gafni, O., Nachmani, E., and Taigman, Y. (2023) · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Zhang, L. and Agrawala, M. (2023) · 2023
Closest in time.