Fetching the paper…
Reading the bibliography…
Recent advances in diffusion models have significantly enhanced the quality of image synthesis, yet they have also introduced serious safety concerns, particularly the generation of Not Safe for Work (NSFW) content.
Auto-encoding variational bayes
Kingma, D. P · 2013
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
nsfwdata, 2020
Kim, A · 2020
Earlier work this paper cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Meng, C., He, Y., Song, Y., Song, J., Wu, J., Zhu, J.-Y., and Ermon, S · 2021
Earlier work this paper cites.
On generating transferable targeted perturbations
Naseer, M., Khan, S., Hayat, M., Khan, F. S., and Porikli, F · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Earlier work this paper cites.
Safety checker nested in stable diffusion., 2022
CompVis · 2022
Earlier work this paper cites.
Diffusion models for adversarial purification
Nie, W., Guo, B., Huang, Y., Xiao, C., Vahdat, A., and Anandkumar, A · 2022
Earlier work this paper cites.
Hierarchical text-conditional image generation with clip latents
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Earlier work this paper cites.
Can machines help us answering question 16 in datasheets, and in turn reflecting on inappropriate content?
Schramowski, P., Tauchmann, C., and Kersting, K · 2022
Earlier work this paper cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al · 2022
Earlier work this paper cites.
Extracting latent steering vectors from pretrained language models
Subramani, N., Suresh, N., and Peters, M. E · 2022
Cited alongside, same era.
https://pypi.org/project/nudenet/
Nudenet, 2023 · 2023
Cited alongside, same era.
Detecting language model attacks with perplexity
Alon, G. and Kamfonas, M · 2023
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Brooks, T., Holynski, A., and Efros, A. A · 2023
Cited alongside, same era.
Erasing concepts from diffusion models
Gandikota, R., Materzynska, J., Fiotto-Kaufman, J., and Bau, D · 2023
Cited alongside, same era.
Character as pixels: A controllable prompt adversarial attacking framework for black-box text guided image generation models
Topiq: A top-down approach from semantics to distortions for image quality assessment
Chen, C., Mo, J., Hou, J., Wu, H., Liao, L., Sun, W., Yan, Q., and Lin, W · 2024
Closest in time.
Scaling rectified flow transformers for high-resolution image synthesis
Esser, P., Kulal, S., Blattmann, A., Entezari, R., Müller, J., Saini, H., Levi, Y., Lorenz, D., Sauer, A., Boesel, F., et al · 2024
Closest in time.
Unified concept editing in diffusion models
Gandikota, R., Orgad, H., Belinkov, Y., Materzyńska, J., and Bau, D · 2024
Closest in time.
Adversarial perturbations cannot reliably protect artists from generative ai
Hönig, R., Rando, J., Carlini, N., and Tramèr, F · 2024
Closest in time.
Latent guard: a safety framework for text-to-image generation
Liu, R., Khakzar, A., Gu, J., Chen, Q., Torr, P., and Pizzati, F · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kou, Z., Pei, S., Tian, Y., and Zhang, X · 2023
Cited alongside, same era.
Visual instruction inversion: Image editing via visual prompting
Nguyen, T., Li, Y., Ojha, U., and Lee, Y. J · 2023
Cited alongside, same era.
Zero-shot image-to-image translation
Parmar, G., Kumar Singh, K., Zhang, R., Li, Y., Lu, J., and Zhu, J.-Y · 2023
Cited alongside, same era.
Safe latent diffusion: Mitigating inappropriate degeneration in diffusion models
Schramowski, P., Brack, M., Deiseroth, B., and Kersting, K · 2023
Cited alongside, same era.
Glaze: Protecting artists from style mimicry by { \{ Text-to-Image } \} models
Shan, S., Cryan, J., Wenger, E., Zheng, H., Hanocka, R., and Zhao, B. Y · 2023
Cited alongside, same era.
Ring-a-bell! how reliable are concept removal methods for diffusion models?
Tsai, Y.-L., Hsu, C.-Y., Xie, C., Lin, C.-H., Chen, J.-Y., Li, B., Chen, P.-Y., Yu, C.-M., and Huang, C.-Y · 2023
Cited alongside, same era.
Adding conditional control to text-to-image diffusion models
Zhang, L., Rao, A., and Agrawala, M · 2023
Cited alongside, same era.
Jailbreaking prompt attack: A controllable adversarial attack against diffusion models
Ma, J., Cao, A., Xiao, Z., Zhang, J., Ye, C., and Zhao, J · 2024
Closest in time.
Chatgpt, 2024
OpenAI · 2024
Closest in time.
Robust concept erasure using task vectors
Pham, M., Marshall, K. O., Hegde, C., and Cohen, N · 2024
Closest in time.
Attacks and defenses for generative diffusion models: A comprehensive survey
Truong, V. T., Dang, L. B., and Le, L. B · 2024
Closest in time.
Universal prompt optimizer for safe text-to-image generation
Wu, Z., Gao, H., Wang, Y., Zhang, X., and Wang, S · 2024
Closest in time.
Sneakyprompt: Jailbreaking text-to-image generative models
Yang, Y., Hui, B., Yuan, H., Gong, N., and Cao, Y · 2024
Closest in time.
Face-adapter for pre-trained diffusion models with fine-grained id and attribute control
Han, Y., Zhu, J., He, K., Chen, X., Ge, Y., Li, W., Li, X., Zhang, J., Wang, C., and Liu, Y · 2025
Closest in time.