Fetching the paper…
Reading the bibliography…
In recent years, the fashion industry has increasingly adopted AI technologies to enhance customer experience, driven by the proliferation of e-commerce platforms and virtual applications.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE Trans. Image Processing , vol. 13, no. 4, pp. 600–612, 2004
2004
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in NeurIPS , 2014
2014
Earlier work this paper cites.
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli, “Deep Unsupervised Learning using Nonequilibrium Thermodynamics,” in ICML , 2015
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-Net: Convolutional Networks for Biomedical Image Segmentation,” in MICCAI , 2015
2015
Earlier work this paper cites.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, G. Klambauer, and S. Hochreiter, “GANs trained by a two time-scale update rule converge to a Nash equilibrium,” in NeurIPS , 2017
2017
Earlier work this paper cites.
X. Han, Z. Wu, Z. Wu, R. Yu, and L. S. Davis, “VITON: An Image-Based Virtual Try-On Network,” in CVPR , 2018
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The Unreasonable Effectiveness of Deep Features as a Perceptual Metric,” in CVPR , 2018
2018
Earlier work this paper cites.
M. Bińkowski, D. J. Sutherland, M. Arbel, and A. Gretton, “Demystifying MMD GANs,” in ICLR , 2018
2018
Earlier work this paper cites.
I. Loshchilov and F. Hutter, “Decoupled Weight Decay Regularization,” in ICLR , 2019
2019
Earlier work this paper cites.
M. Fincato, F. Landi, M. Cornia, F. Cesari, and R. Cucchiara, “VITON-GT: An Image-based Virtual Try-On Model with Geometric Transformations,” in ICPR , 2020
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising Diffusion Probabilistic Models,” in NeurIPS , 2020
2020
Earlier work this paper cites.
S. Choi, S. Park, M. Lee, and J. Choo, “VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization,” in CVPR , 2021
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin et al. , “Learning Transferable Visual Models From Natural Language Supervision,” in ICML , 2021
2021
Earlier work this paper cites.
P. Dhariwal and A. Nichol, “Diffusion Models Beat GANs on Image Synthesis,” in NeurIPS , 2021
2021
Earlier work this paper cites.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby, “An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale,” in ICLR , 2021
2021
Earlier work this paper cites.
D. Morelli, M. Fincato, M. Cornia, F. Landi, F. Cesari, and R. Cucchiara, “Dress Code: High-Resolution Multi-Category Virtual Try-On,” in ECCV , 2022
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-Resolution Image Synthesis With Latent Diffusion Models,” in CVPR , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
A. Blattmann, R. Rombach, K. Oktay, J. Müller, and B. Ommer, “Retrieval-Augmented Diffusion Models,” in NeurIPS , 2022
2022
Earlier work this paper cites.
D. Morelli, A. Baldrati, G. Cartella, M. Cornia, M. Bertini, and R. Cucchiara, “LaDI-VTON: Latent Diffusion Textual-Inversion Enhanced Virtual Try-On,” in ACM Multimedia , 2023
2023
Cited alongside, same era.
A. Baldrati, D. Morelli, G. Cartella, M. Cornia, M. Bertini, and R. Cucchiara, “Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing,” in ICCV , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
R. Gal, Y. Alaluf, Y. Atzmon, O. Patashnik, A. H. Bermano, G. Chechik, and D. Cohen-Or, “An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion,” in ICLR , 2023
2023
Cited alongside, same era.
L. Zhang and M. Agrawala, “Adding Conditional Control to Text-to-Image Diffusion Models,” in ICCV , 2023
2023
Later among the works it cites.
M. Cherti, R. Beaumont, R. Wightman, M. Wortsman, G. Ilharco, C. Gordon, C. Schuhmann et al. , “Reproducible Scaling Laws for Contrastive Language-Image Learning,” in CVPR , 2023
2023
Later among the works it cites.
O. Avrahami, T. Hayes, O. Gafni, S. Gupta, Y. Taigman, D. Parikh, D. Lischinski, O. Fried, and X. Yin, “SpaText: Spatio-Textual Representation for Controllable Image Generation,” in CVPR , 2023
2023
Later among the works it cites.
P. Esser, S. Kulal, A. Blattmann, R. Entezari, J. Müller, H. Saini, Y. Levi, D. Lorenz, A. Sauer, F. Boesel et al. , “Scaling Rectified Flow Transformers for High-Resolution Image Synthesis,” in ICML , 2024
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Baldrati, L. Agnolucci, M. Bertini, and A. Del Bimbo, “Zero-Shot Composed Image Retrieval with Textual Inversion,” ICCV , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
J. Gou, S. Sun, J. Zhang, J. Si, C. Qian, and L. Zhang, “Taming the Power of Diffusion Models for High-Quality Virtual Try-On with Appearance Flow,” in ACM Multimedia , 2023
2023
Cited alongside, same era.
C.-Y. Chen, Y.-C. Chen, H.-H. Shuai, and W.-H. Cheng, “Size Does Matter: Size-aware Virtual Try-on via Clothing-oriented Transformation Try-on Network,” in ICCV , 2023
2023
Cited alongside, same era.
L. Zhu, D. Yang, T. Zhu, F. Reda, W. Chan, C. Saharia, M. Norouzi, and I. Kemelmacher-Shlizerman, “TryOnDiffusion: A Tale of Two UNets,” in CVPR , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
D. Cheng, S. Huang, J. Bi, Y. Zhan, J. Liu, Y. Wang, H. Sun, F. Wei, D. Deng, and Q. Zhang, “UPRISE: Universal Prompt Retrieval for Improving Zero-Shot Evaluation,” in EMNLP , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Choi, S. Kwak, K. Lee, H. Choi, and J. Shin, “Improving Diffusion Models for Authentic Virtual Try-on in the Wild,” in ECCV , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
T. Wang and M. Ye, “TexFit: Text-Driven Fashion Image Editing with Diffusion Models,” in AAAI , 2024
2024
Later among the works it cites.
A. Asai, Z. Wu, Y. Wang, A. Sil, and H. Hajishirzi, “Self-RAG: Learning to retrieve, generate, and critique through self-reflection,” in ICLR , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
S. Sarto, M. Cornia, L. Baraldi, A. Nicolosi, and R. Cucchiara, “Towards Retrieval-Augmented Architectures for Image Captioning,” ACM TOMM , vol. 20, no. 8, pp. 1–22, 2024
2024
Later among the works it cites.
A. Hertz, A. Voynov, S. Fruchter, and D. Cohen-Or, “Style Aligned Image Generation via Shared Attention,” in CVPR , 2024
2024
Later among the works it cites.
Y. Alaluf, D. Garibi, O. Patashnik, H. Averbuch-Elor, and D. Cohen-Or, “Cross-Image Attention for Zero-Shot Appearance Transfer,” in ACM SIGGRAPH , 2024
2024
Later among the works it cites.
Z. Chong, X. Dong, H. Li, S. Zhang, W. Zhang, X. Zhang, H. Zhao, D. Jiang, and X. Liang, “CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models,” in ICLR , 2025
2025
Closest in time.
F. Cocchi, N. Moratelli, M. Cornia, L. Baraldi, and R. Cucchiara, “Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering,” in CVPR , 2025
2025
Closest in time.
D. Caffagni, S. Sarto, M. Cornia, L. Baraldi, and R. Cucchiara, “Recurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval,” in CVPR , 2025
2025
Closest in time.