Fetching the paper…
Reading the bibliography…
The fashion industry is increasingly leveraging computer vision and deep learning technologies to enhance online shopping experiences and operational efficiencies.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, A.C. Bovik, H.R. Sheikh, and E.P. Simoncelli. 2004 · 2003
Earlier work this paper cites.
Image Quality Assessment: Unifying Structure and Texture Similarity
Keyan Ding, Kede Ma, Shiqi Wang, and Eero P. Simoncelli. 2020 · 2004
Earlier work this paper cites.
Denoising Diffusion Implicit Models
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2020 · 2010
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Diederik P Kingma and Max Welling. 2013 · 2013
Earlier work this paper cites.
Deep Human Parsing with Active Template Regression
Xiaodan Liang, Si Liu, Xiaohui Shen, Jianchao Yang, Luoqi Liu, Jian Dong, Liang Lin, and Shuicheng Yan. 2015 · 2015
Earlier work this paper cites.
U-Net: Convolutional networks for biomedical image segmentation. In International Conference on Medical Image Computing and Computer-Assisted Intervention . Springer, 234–241
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
Image-to-Image Translation with Conditional Adversarial Networks. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 5967–5976
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros. 2017b · 2017
Earlier work this paper cites.
Decoupled Weight Decay Regularization
Ilya Loshchilov and Frank Hutter. 2017 · 2017
Earlier work this paper cites.
Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. In Proceedings of the IEEE International Conference on Computer Vision (ICCV)
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros. 2017 · 2017
Earlier work this paper cites.
Densepose: Dense human pose estimation in the wild
Rıza Alp G¨ uler, Natalia Neverova, and Iasonas Kokkinos. 2018 · 2018
Earlier work this paper cites.
The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. In 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition . 586–595
Richard Zhang, Phillip Isola, Alexei A. Efros, Eli Shechtman, and Oliver Wang. 2018 · 2018
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models. In Advances in Neural Information Processing Systems , Vol. 33. 6840–6851
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
pytorch-fid: FID Score for PyTorch
Maximilian Seitzer. 2020 · 2020
Earlier work this paper cites.
TileGAN: category-oriented attention-based high-quality tiled clothes generation from dressed person
Wei Zeng, Mingbo Zhao, Yuan Gao, Liuwu Li, and Wenyin Liu. 2020 · 2020
Earlier work this paper cites.
Blended Diffusion for Complex Object Insertion
Omer Avrahami, Bar Lavi, and Daniel Cohen-Or. 2021 · 2021
Earlier work this paper cites.
Mikołaj Bińkowski, Danica J. Sutherland, Michael Arbel, and Arthur Gretton. 2021 · 2021
Earlier work this paper cites.
Fashion Meets Computer Vision: A Survey
Wen-Huang Cheng, Sijie Song, Chieh-Yun Chen, Shintami Chusnul Hidayati, and Jiaying Liu. 2021 · 2021
Cited alongside, same era.
Diffusion Models Beat GANs on Image Synthesis. In Advances in Neural Information Processing Systems , Vol. 34. 8780–8794
Prafulla Dhariwal and Alex Nichol. 2021 · 2021
Cited alongside, same era.
DiffusionCLIP: Text-Guided Diffusion Models for Robust Image Manipulation
Giannis Kim, Jaesik Jeong, Junho Ahn, et al · 2021
Cited alongside, same era.
Stochastic Differential Editing for Image Manipulation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Chenlin Meng, Jonathan Ho, Yang Song, Pieter Abbeel, and Jascha Sohl-Dickstein. 2021 · 2021
Cited alongside, same era.
Improved Denoising Diffusion Probabilistic Models
Alexander Quinn Nichol and Prafulla Dhariwal. 2021 · 2021
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S. Sara Mahdavi, Rapha Gontijo Lopes, et al · 2022
Later among the works it cites.
Image Super-Resolution via Iterative Refinement
Chitwan Saharia, Jonathan Ho, William Chan, Tim Salimans, David J Fleet, and Mohammad Norouzi. 2022b · 2022
Later among the works it cites.
Palette: Image-to-Image Diffusion Models
Chitwan Saharia, Jonathan Ho, William Chan, Tim Salimans, David J Fleet, and Mohammad Norouzi. 2022c · 2022
Later among the works it cites.
LAION-5B: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, and Jenia Jitsev. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning Transferable Visual Models From Natural Language Supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021a · 2021
Cited alongside, same era.
Viton-hd: High-resolution virtual try-on via misalignment-aware normalization.. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Seunghwan Choi, Sunghyun Park, Minsoo Lee, and Jaegul Choo. 2021 · 2021
Cited alongside, same era.
SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers
Enze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar, Jose M. Alvarez, and Ping Luo. 2021 · 2021
Cited alongside, same era.
eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers
Yash Balaji, Mathilde Caron, Alaaeldin El-Nouby, Gabriel Synnaeve, Armand Joulin, and Ishan Misra. 2022 · 2022
Cited alongside, same era.
InstructPix2Pix: Editing Images with Instructions Using Diffusion Models
Tim Brooks, Aleksander Holynski, and Alexei A Efros. 2022 · 2022
Cited alongside, same era.
Imagic: Text-based Real Image Editing with Diffusion Models
Bahjat Kawar, Aparna Vinod, Micha Elad, et al · 2022
Cited alongside, same era.
Palette: Image-to-Image Diffusion Models for High Fidelity and Diverse Colorization and Inpainting. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 11523–11532
Andreas Lugmayr, Martin Danelljan, and Radu Timofte. 2022 · 2022
Cited alongside, same era.
Binxin Yang, Shuyang Gu, Bo Zhang, Ting Zhang, Xuejin Chen, Xiaoyan Sun, Dong Chen, and Fang Wen. 2022 · 2022
Later among the works it cites.
AnyDoor: Zero-Shot Personalized Image Generation and Manipulation
Wenhu Chen, Hexiang Hu, Yandong Li, Nataniel Ruiz, Xuhui Jia, Ming-Wei Chang, and William W Cohen. 2023 · 2023
Later among the works it cites.
Computational Technologies for Fashion Recommendation: A Survey
Yujuan Ding, Zhihui Lai, P. Y. Mok, and Tat-Seng Chua. 2023 · 2023
Later among the works it cites.
Dongxu Li, Junnan Li, and Steven CH Hoi. 2023 · 2023
Later among the works it cites.
Subject-Diffusion: Open-Domain Personalized Text-to-Image Generation without Test-Time Fine-Tuning
Yifan Liu, Yuheng Zhang, Yulun Zhang, and Yunfu Zhang. 2023 · 2023
Later among the works it cites.
Sigmoid Loss for Language Image Pre-Training
Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, and Lucas Beyer. 2023 · 2023
Later among the works it cites.
Adding Conditional Control to Text-to-Image Diffusion Models
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala. 2023a · 2023
Later among the works it cites.
FastComposer: Tuning-Free Multi-Subject Image Generation with Localized Attention
Yuheng Zhang, Zhaoyang Zhang, Yubin Zhang, Yifan Zhang, Yulun Zhang, and Yunfu Zhang. 2023b · 2023
Later among the works it cites.
CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models
Zheng Chong, Xiao Dong, Haoxiang Li, Shiyue Zhang, Wenqing Zhang, Xujie Zhang, Hanqing Zhao, and Xiaodan Liang. 2024 · 2024
Closest in time.
MS-Diffusion: Multi-subject Zero-shot Image Personalization with Layout Guidance
X Wang, Siming Fu, Qihan Huang, Wanggui He, and Hao Jiang. 2024 · 2024
Closest in time.
Good Seed Makes a Good Crop: Discovering Secret Seeds in Text-to-Image Diffusion Models
Katherine Xu, Lingzhi Zhang, and Jianbo Shi. 2024 · 2024
Closest in time.