Fetching the paper…
Reading the bibliography…
Multimodal machine learning, especially text-to-image models like Stable Diffusion and DALL-E 3, has gained significance for transforming text into detailed images.
Improving image captioning with better use of captions
Shi, Z.; Zhou, X.; Qiu, X.; and Zhu, X. 2020 · 2006
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Earlier work this paper cites.
Membership inference attacks against machine learning models
Shokri, R.; Stronati, M.; Song, C.; and Shmatikov, V. 2017 · 2017
Earlier work this paper cites.
The secret sharer: Evaluating and testing unintended memorization in neural networks
Carlini, N.; Liu, C.; Erlingsson, Ú.; Kos, J.; and Song, D. 2019 · 2019
Earlier work this paper cites.
Does learning require memorization? a short tale about a long tail
Feldman, V. 2020 · 2020
Earlier work this paper cites.
What neural networks memorize and why: Discovering the long tail via influence estimation
Feldman, V.; and Zhang, C. 2020 · 2020
Earlier work this paper cites.
Extracting training data from large language models
Carlini, N.; Tramer, F.; Wallace, E.; Jagielski, M.; Herbert-Voss, A.; Lee, K.; Roberts, A.; Brown, T.; Song, D.; Erlingsson, U.; et al. 2021 · 2021
Earlier work this paper cites.
Taming transformers for high-resolution image synthesis
Esser, P.; Rombach, R.; and Ommer, B. 2021 · 2021
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; McGrew, B.; Sutskever, I.; and Chen, M. 2021 · 2021
Cited alongside, same era.
Laion-400m: Open dataset of clip-filtered 400 million image-text pairs
Schuhmann, C.; Vencu, R.; Beaumont, R.; Kaczmarczyk, R.; Mullis, C.; Katta, A.; Coombes, T.; Jitsev, J.; and Komatsuzaki, A. 2021 · 2021
Cited alongside, same era.
Reconstructing training data with informed adversaries
Balle, B.; Cherubin, G.; and Hayes, J. 2022 · 2022
Cited alongside, same era.
Vector quantized diffusion model for text-to-image synthesis
Gu, S.; Chen, D.; Bao, J.; Wen, F.; Zhang, B.; Chen, D.; Yuan, L.; and Guo, B. 2022 · 2022
Cited alongside, same era.
Deduplicating training data mitigates privacy risks in language models
Kandpal, N.; Wallace, E.; and Raffel, C. 2022 · 2022
Cited alongside, same era.
A self-supervised descriptor for image copy detection
Pizzi, E.; Roy, S. D.; Ravindra, S. N.; Goyal, P.; and Douze, M. 2022 · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, C.; Beaumont, R.; Vencu, R.; Gordon, C.; Wightman, R.; Cherti, M.; Coombes, T.; Katta, A.; Mullis, C.; Wortsman, M.; et al. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stylet2i: Toward compositional and high-fidelity text-to-image synthesis
Li, Z.; Min, M. R.; Li, K.; and Xu, C. 2022 · 2022
Cited alongside, same era.
Midjourney.com
Midjourney. 2022 · 2022
Cited alongside, same era.
An empirical analysis of memorization in fine-tuned autoregressive language models
Mireshghallah, F.; Uniyal, A.; Wang, T.; Evans, D. K.; and Berg-Kirkpatrick, T. 2022 · 2022
Cited alongside, same era.
Diffusion art or digital forgery? investigating data replication in diffusion models
Somepalli, G.; Singla, V.; Goldblum, M.; Geiping, J.; and Goldstein, T. 2023a
Cited in the paper.
Understanding and Mitigating Copying in Diffusion Models
Somepalli, G.; Singla, V.; Goldblum, M.; Geiping, J.; and Goldstein, T. 2023b
Cited in the paper.
Extracting training data from diffusion models
Carlini, N.; Hayes, J.; Nasr, M.; Jagielski, M.; Sehwag, V.; Tramer, F.; Balle, B.; Ippolito, D.; and Wallace, E. 2023 · 2023
Closest in time.
Forget-me-not: Learning to forget in text-to-image diffusion models
Zhang, E.; Wang, K.; Xu, X.; Wang, Z.; and Shi, H. 2023 · 2023
Closest in time.