Fetching the paper…
Reading the bibliography…
Text-driven person image generation is an emerging and challenging task in cross-modality image generation.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Multiscale structural similarity for image quality assessment
Wang, Z. and Simoncelli, E. P. (2003) · 2003
Earlier work this paper cites.
Pytorch distributed: Experiences on accelerating data parallel training
Li, S., Zhao, Y., Varma, R., Salpekar, O., Noordhuis, P., Li, T., Paszke, A., Smith, J., Vaughan, B., Damania, P., et al. (2020) · 2006
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S. (2020a) · 2010
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Song, Y., Sohl-Dickstein, J., Kingma, D. P., Kumar, A., Ermon, S., and Poole, B. (2020b) · 2011
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. (2014) · 2014
Earlier work this paper cites.
Weston, J., Chopra, S., and Bordes, A. (2014) · 2014
Earlier work this paper cites.
Tutorial on variational autoencoders
Doersch, C. (2016) · 2016
Earlier work this paper cites.
Deepfashion: Powering robust clothes recognition and retrieval with rich annotations
Liu, Z., Luo, P., Qiu, S., Wang, X., and Tang, X. (2016) · 2016
Earlier work this paper cites.
Realtime multi-person 2d pose estimation using part affinity fields
Cao, Z., Simon, T., Wei, S.-E., and Sheikh, Y. (2017) · 2017
Earlier work this paper cites.
Attend to you: Personalized image captioning with context sequence memory networks
Chunseong Park, C., Kim, B., and Kim, G. (2017) · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., and Nessler, B. (2017) · 2017
Earlier work this paper cites.
Soft-gated warping-gan for pose-guided person image synthesis
Dong, H., Liang, X., Gong, K., Lai, H., Zhu, J., and Yin, J. (2018) · 2018
Earlier work this paper cites.
A variational u-net for conditional appearance and shape generation
Esser, P., Sutter, E., and Ommer, B. (2018) · 2018
Earlier work this paper cites.
Disentangled person image generation
Ma, L., Sun, Q., Georgoulis, S., Van Gool, L., and Schiele (2018) · 2018
Earlier work this paper cites.
Unsupervised person image synthesis in arbitrary poses
Pumarola, A., Agudo, A., Sanfeliu, A., and Moreno-Noguer, F. (2018) · 2018
Earlier work this paper cites.
Person re-identification with deep similarity-guided graph neural network
Shen, Y., Li, H., Yi, S., Chen, D., and Wang, X. (2018) · 2018
Earlier work this paper cites.
Deformable gans for pose-based human image generation
Siarohin, A., Sangineto, E., and Lathuiliere, S. (2018) · 2018
Earlier work this paper cites.
High-resolution image synthesis and semantic manipulation with conditional gans
Wang, T.-C., Liu, M.-Y., Zhu, J.-Y., Tao, A., Kautz, J., and Catanzaro, B. (2018) · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Zhang, R., Isola, P., Efros, A. A., Shechtman, E., and Wang, O. (2018) · 2018
Earlier work this paper cites.
Clothflow: A flow-based model for clothed person generation
Han, X., Hu, X., Huang, W., and Scott, M. R. (2019) · 2019
Earlier work this paper cites.
Dense intrinsic appearance flow for human pose transfer
Li, Y., Huang, C., and Loy, C. C. (2019) · 2019
Cited alongside, same era.
Semantic image synthesis with spatially-adaptive normalization
Park, T. and Liu, M.-Y. (2019) · 2019
Cited alongside, same era.
Generative modeling by estimating gradients of the data distribution
Song, Y. and Ermon, S. (2019) · 2019
Cited alongside, same era.
Dm-gan: Dynamic memory generative adversarial networks for text-to-image synthesis
Zhu, M. and Pan, P. (2019) · 2019
Cited alongside, same era.
Progressive pose attention transfer for person image generation
Zhu, Z., Huang, T., Shi, B., Yu, M., Wang, B., and Bai, X. (2019) · 2019
Cited alongside, same era.
Generative adversarial networks
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2020) · 2020
Lafite: Towards language-free training for text-to-image generation
Zhou, Y., Zhang, R., Chen, C., Li, C., Tensmeyer, C., Yu, T., Gu, J., Xu, J., and Sun, T. (2021) · 2021
Later among the works it cites.
Blended diffusion for text-driven editing of natural images
Avrahami, O., Lischinski, D., and Fried, O. (2022) · 2022
Closest in time.
Retrieval-augmented diffusion models
Blattmann, A., Rombach, R., Oktay, K., and Ommer, B. (2022) · 2022
Closest in time.
Pose guided multi-person image generation from text
Cheong, S. Y., Mustafa, A., and Gilbert, A. (2022) · 2022
Closest in time.
Stylegan-human: A data-centric odyssey of human generation
Fu, J., Li, S., Jiang, Y., Lin, K.-Y., Qian, C., Loy, C. C., Wu, W., and Liu, Z. (2022) · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P. (2020) · 2020
Cited alongside, same era.
Memory aggregation networks for efficient interactive video object segmentation
Miao, J. and Wei, Y. (2020) · 2020
Cited alongside, same era.
Pose with style: Detail-preserving pose-guided image synthesis with conditional stylegan
Albahar, B., Lu, J., Yang, J., Shu, Z., and Shechtman (2021) · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Dhariwal, P. and Nichol, A. (2021) · 2021
Cited alongside, same era.
Memory oriented transfer learning for semi-supervised image deraining
Huang, H., Yu, A., and He, R. (2021) · 2021
Cited alongside, same era.
Tryongan: Body-aware try-on via layered interpolation
Lewis, K. M., Varadharajan, S., and Kemelmacher-Shlizerman, I. (2021) · 2021
Cited alongside, same era.
Make-a-scene: Scene-based text-to-image generation with human priors
Gafni, O., Polyak, A., Ashual, O., Sheynin, S., Parikh, D., and Taigman, Y. (2022) · 2022
Closest in time.
Vector quantized diffusion model for text-to-image synthesis
Gu, S., Chen, D., Bao, J., Wen, F., Zhang, B., Chen, D., and Yuan, L. (2022) · 2022
Closest in time.
Style-based global appearance flow for virtual try-on
He, S., Song, Y.-Z., and Xiang, T. (2022) · 2022
Closest in time.
Classifier-free diffusion guidance
Ho, J. and Salimans, T. (2022) · 2022
Closest in time.
Imagic: Text-based real image editing with diffusion models
Kawar, B., Zada, S., Lang, O., Tov, O., Chang, H., Dekel, T., Mosseri, I., and Irani, M. (2022) · 2022
Closest in time.
Diffusionclip: Text-guided diffusion models for robust image manipulation
Kim, G. and Kwon, T. (2022) · 2022
Closest in time.
Text to image generation with semantic-spatial aware gan
Liao, W., Hu, K., Yang, M. Y., and Rosenhahn, B. (2022) · 2022
Closest in time.
Verbal-person nets: Pose-guided multi-granularity language-to-person generation
Liu, D., Wu, L., Zheng, F., Liu, L., and Wang, M. (2022) · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M. (2022) · 2022
Closest in time.
Neural texture extraction and distribution for controllable person image synthesis
Ren, Y., Fan, X., Li, G., Liu, S., and Li, T. H. (2022) · 2022
Closest in time.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B. (2022) · 2022
Closest in time.
Knn-diffusion: Image generation via large-scale retrieval
Sheynin, S., Ashual, O., Polyak, A., Singer, U., Gafni, O., Nachmani, E., and Taigman, Y. (2022) · 2022
Closest in time.
Anyface: Free-style text-to-face synthesis and manipulation
Sun, J., Deng, Q., Li, Q., Sun, M., Ren, M., and Sun, Z. (2022) · 2022
Closest in time.
Exploring dual-task correlation for pose guided person image generation
Zhang, P. and Yang, L. (2022) · 2022
Closest in time.
Cross attention based style distribution for controllable person image synthesis
Zhou, X., Yin, M., Chen, X., Sun, L., Gao, C., and Li, Q. (2022) · 2022
Closest in time.
Styleclip: Text-driven manipulation of stylegan imagery
Patashnik, O., Wu, Z., Shechtman, E., Cohen-Or, D., and Lischinski, D. (2021) · 2094
Closest in time.