Fetching the paper…
Reading the bibliography…
Latest advances have achieved realistic virtual try-on (VTON) through localized garment inpainting using latent diffusion models, significantly enhancing consumers' online shopping experience.
Auto-encoding variational bayes
Kingma, D. P.; and Welling, M. 2013 · 2013
Earlier work this paper cites.
Realtime multi-person 2d pose estimation using part affinity fields
Cao, Z.; Simon, T.; Wei, S.-E.; and Sheikh, Y. 2017 · 2017
Earlier work this paper cites.
Generative adversarial networks: An overview
Creswell, A.; White, T.; Dumoulin, V.; Arulkumaran, K.; Sengupta, B.; and Bharath, A. A. 2018 · 2018
Earlier work this paper cites.
Densepose: Dense human pose estimation in the wild
Güler, R. A.; Neverova, N.; and Kokkinos, I. 2018 · 2018
Earlier work this paper cites.
Viton: An image-based virtual try-on network
Han, X.; Wu, Z.; Wu, Z.; Yu, R.; and Davis, L. S. 2018 · 2018
Earlier work this paper cites.
Clothflow: A flow-based model for clothed person generation
Han, X.; Hu, X.; Huang, W.; and Scott, M. R. 2019 · 2019
Earlier work this paper cites.
Sc-fegan: Face editing generative adversarial network with user’s sketch and color
Jo, Y.; and Park, J. 2019 · 2019
Earlier work this paper cites.
Fashion editing with adversarial parsing learning
Dong, H.; Liang, X.; Zhang, Y.; Zhang, X.; Shen, X.; Xie, Z.; Wu, B.; and Yin, J. 2020 · 2020
Earlier work this paper cites.
Do not mask what you do not need to mask: a parser-free virtual try-on
Issenhuth, T.; Mary, J.; and Calauzenes, C. 2020 · 2020
Earlier work this paper cites.
Maskgan: Towards diverse and interactive facial image manipulation
Lee, C.-H.; Liu, Z.; Wu, L.; and Luo, P. 2020 · 2020
Earlier work this paper cites.
Self-correction for human parsing
Li, P.; Xu, Y.; Wei, Y.; and Yang, Y. 2020 · 2020
Earlier work this paper cites.
MGCM: Multi-modal generative compatibility modeling for clothing matching
Liu, J.; Song, X.; Chen, Z.; and Ma, J. 2020 · 2020
Earlier work this paper cites.
Viton-hd: High-resolution virtual try-on via misalignment-aware normalization
Choi, S.; Park, S.; Lee, M.; and Choo, J. 2021 · 2021
Earlier work this paper cites.
Parser-free virtual try-on via distilling appearance flows
Ge, Y.; Song, Y.; Zhang, R.; Ge, C.; Liu, W.; and Luo, P. 2021 · 2021
Earlier work this paper cites.
Tryongan: Body-aware try-on via layered interpolation
Lewis, K. M.; Varadharajan, S.; and Kemelmacher-Shlizerman, I. 2021 · 2021
Earlier work this paper cites.
Toward accurate and realistic outfits visualization with attention to details
Li, K.; Chong, M. J.; Zhang, J.; and Liu, J. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J.; and Salimans, T. 2022 · 2022
Cited alongside, same era.
High-resolution virtual try-on with misalignment and occlusion-handled conditions
Lee, S.; Gu, G.; Park, S.; Choi, S.; and Choo, J. 2022 · 2022
Cited alongside, same era.
Dress code: High-resolution multi-category virtual try-on
Morelli, D.; Fincato, M.; Cornia, M.; Landi, F.; Cesari, F.; and Cucchiara, R. 2022 · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Cited alongside, same era.
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
Xu, X.; Wang, Z.; Zhang, E. J.; Wang, K.; and Shi, H. 2023 · 2023
Later among the works it cites.
IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Ye, H.; Zhang, J.; Liu, S.; Han, X.; and Yang, W. 2023 · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Zhang, L.; Rao, A.; and Agrawala, M. 2023 · 2023
Later among the works it cites.
Tryondiffusion: A tale of two unets
Zhu, L.; Yang, D.; Zhu, T.; Reda, F.; Chan, W.; Saharia, C.; Norouzi, M.; and Kemelmacher-Shlizerman, I. 2023 · 2023
Later among the works it cites.
Magic Clothing: Controllable Garment-Driven Image Synthesis
Chen, W.; Gu, T.; Xu, Y.; and Chen, C. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Cited alongside, same era.
Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Bai, J.; Bai, S.; Yang, S.; Wang, S.; Tan, S.; Wang, P.; Lin, J.; Zhou, C.; and Zhou, J. 2023 · 2023
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Brooks, T.; Holynski, A.; and Efros, A. A. 2023 · 2023
Cited alongside, same era.
Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing
Cao, M.; Wang, X.; Qi, Z.; Shan, Y.; Qie, X.; and Zheng, Y. 2023 · 2023
Cited alongside, same era.
Taming the power of diffusion models for high-quality virtual try-on with appearance flow
Gou, J.; Sun, S.; Zhang, J.; Si, J.; Qian, C.; and Zhang, L. 2023 · 2023
Cited alongside, same era.
BLIP-Diffusion: Pre-trained Subject Representation for Controllable Text-to-Image Generation and Editing
Li, D.; Li, J.; and Hoi, S. C. H. 2023 · 2023
Cited alongside, same era.
Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Li, J.; Li, D.; Savarese, S.; and Hoi, S. 2023 · 2023
Cited alongside, same era.
Choi, Y.; Kwak, S.; Lee, K.; Choi, H.; and Shin, J. 2024 · 2024
Closest in time.
Dong, X.; Zhang, P.; Zang, Y.; Cao, Y.; Wang, B.; Ouyang, L.; Wei, X.; Zhang, S.; Duan, H.; Cao, M.; Zhang, W.; Li, Y.; Yan, H.; Gao, Y.; Zhang, X.; Li, W.; Li, J.; Chen, K.; He, C.; Zhang, X.; Qiao, Y.; Lin, D.; and Wang, J. 2024 · 2024
Closest in time.
Stableviton: Learning semantic correspondence with latent diffusion model for virtual try-on
Kim, J.; Gu, G.; Park, M.; Park, S.; and Choo, J. 2024 · 2024
Closest in time.
Improved baselines with visual instruction tuning
Liu, H.; Li, C.; Li, Y.; and Lee, Y. J. 2024 · 2024
Closest in time.
T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models
Mou, C.; Wang, X.; Xie, L.; Wu, Y.; Zhang, J.; Qi, Z.; and Shan, Y. 2024 · 2024
Closest in time.
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
Shen, F.; Ye, H.; Liu, S.; Zhang, J.; Wang, C.; Han, X.; and Yang, W. 2024 · 2024
Closest in time.
Ensembling Diffusion Models via Adaptive Feature Aggregation
Wang, C.; Tian, K.; Guan, Y.; Zhang, J.; Jiang, Z.; Shen, F.; Han, X.; Gu, Q.; and Yang, W. 2024 · 2024
Closest in time.
Texture-Preserving Diffusion Models for High-Fidelity Virtual Try-On
Yang, X.; Ding, C.; Hong, Z.; Huang, J.; Tao, J.; and Xu, X. 2024 · 2024
Closest in time.
CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model
Zeng, J.; Song, D.; Nie, W.; Tian, H.; Wang, T.; and Liu, A.-A. 2024 · 2024
Closest in time.
Uni-controlnet: All-in-one control to text-to-image diffusion models
Zhao, S.; Chen, D.; Chen, Y.-C.; Bao, J.; Hao, S.; Yuan, L.; and Wong, K.-Y. K. 2024 · 2024
Closest in time.