Fetching the paper…
Reading the bibliography…
The rapid evolution of multimodal foundation models has led to significant advancements in cross-modal understanding and generation across diverse modalities, including text, images, audio, and video.
S. Sivanandam, S. Deepa, S. Sivanandam, and S. Deepa, Genetic algorithms . Springer, 2008
2008
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Proceedings of the European Conference on Computer Vision , 2014
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Proceedings of the European Conference on Computer Vision , 2014
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” in Proceedings of the Advances in Neural Information Processing Systems , 2017
2017
Earlier work this paper cites.
P. Bedapudi, “Nudenet: Neural nets for nudity classification, detection and selective censoring,” 2019
2019
Earlier work this paper cites.
“Image censorship,” https://github.com/lucasxlu/XCloud/ tree/master/research/imgcensor, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Proceedings of the Advances in Neural Information Processing Systems , 2020
2020
Earlier work this paper cites.
L. Hanu and Unitary team, “Detoxify,” Github. https://github.com/unitaryai/detoxify, 2020
2020
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in Proceedings of the International Conference on Machine Learning , 2021
2021
Earlier work this paper cites.
M. Mathew, D. Karatzas, and C. Jawahar, “Docvqa: A dataset for vqa on document images,” in Proceedings of the IEEE Winter Conference on Applications of Computer Vision , 2021
2021
Earlier work this paper cites.
J. Hessel, A. Holtzman, M. Forbes, R. Le Bras, and Y. Choi, “Clipscore: A reference-free evaluation metric for image captioning,” in Proceedings of the Empirical Methods in Natural Language Processing , 2021
2021
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2022
2022
Earlier work this paper cites.
J. Ho, T. Salimans, A. Gritsenko, W. Chan, M. Norouzi, and D. J. Fleet, “Video diffusion models,” in Proceedings of the Advances in Neural Information Processing Systems , 2022
2022
Earlier work this paper cites.
J. Ho and T. Salimans, “Classifier-free diffusion guidance,” arXiv preprint arXiv:2207.12598 , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Rando, D. Paleka, D. Lindner, L. Heim, and F. Tramèr, “Red-teaming the stable diffusion safety filter,” in Proceedings of the Advances in Neural Information Processing Systems Workshops , 2022
2022
Earlier work this paper cites.
P. Schramowski, C. Tauchmann, and K. Kersting, “Can machines help us answering question 16 in datasheets, and in turn reflecting on inappropriate content?” in Proceedings of the ACM Conference on Fairness, Accountability, and Transparency , 2022
2022
Earlier work this paper cites.
A. Masry, D. X. Long, J. Q. Tan, S. Joty, and E. Hoque, “Chartqa: A benchmark for question answering about charts with visual and logical reasoning,” in Proceedings of the Findings of the Association for Computational Linguistics , 2022
2022
Earlier work this paper cites.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
H. Zhang, X. Li, and L. Bing, “Video-llama: An instruction-tuned audio-visual language model for video understanding,” in Proceedings of the Empirical Methods in Natural Language Processing , 2023
2023
Earlier work this paper cites.
J. Christian, “Amazing “jailbreak” bypasses chatgpt’s ethics safeguards,” Futurism, February , 2023
2023
Earlier work this paper cites.
Albert, “Jailbreak chat,” https://www.jailbreakchat.com/ , 2023
2023
Earlier work this paper cites.
H. Li, D. Guo, W. Fan, M. Xu, J. Huang, F. Meng, and Y. Song, “Multi-step jailbreaking privacy attacks on chatgpt,” in Proceedings of the Findings of Empirical Methods in Natural Language Processing , 2023
2023
Earlier work this paper cites.
A. Wei, N. Haghtalab, and J. Steinhardt, “Jailbroken: How does llm safety training fail?” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
W. Dai, J. Li, D. Li, A. M. H. Tiong, J. Zhao, W. Wang, B. Li, P. Fung, and S. C. H. Hoi, “Instructblip: Towards general-purpose vision-language models with instruction tuning,” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
N. Ruiz, Y. Li, V. Jampani, Y. Pritch, M. Rubinstein, and K. Aberman, “Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2023, pp. 22 500–22 510
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
P. Lu, B. Peng, H. Cheng, M. Galley, K. Chang, Y. N. Wu, S. Zhu, and J. Gao, “Chameleon: Plug-and-play compositional reasoning with large language models,” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
W.-L. Chiang, Z. Li, Z. Lin, Y. Sheng, Z. Wu, H. Zhang, L. Zheng, S. Zhuang, Y. Zhuang, J. E. Gonzalez et al. , “Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,” https://vicuna. lmsys. org , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
N. Carlini, M. Nasr, C. A. Choquette-Choo, M. Jagielski, I. Gao, P. W. W. Koh, D. Ippolito, F. Tramer, and L. Schmidt, “Are aligned neural networks adversarially aligned?” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
Y. Zhao, T. Pang, C. Du, X. Yang, C. Li, N.-M. M. Cheung, and M. Lin, “On evaluating adversarial robustness of large vision-language models,” in Proceedings of the Advances in Neural Information Processing Systems , 2023
2023
Earlier work this paper cites.
“Leonardo.Ai,” https://leonardo.ai/2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
R. Gandikota, J. Materzynska, J. Fiotto-Kaufman, and D. Bau, “Erasing concepts from diffusion models,” in Proceedings of the IEEE International Conference on Computer Vision , 2023
2023
Earlier work this paper cites.
C.-P. Huang, K.-P. Chang, C.-T. Tsai, Y.-H. Lai, and Y.-C. F. Wang, “Receler: Reliable concept erasing of text-to-image diffusion models via lightweight erasers,” in Proceedings of the European Conference on Computer Vision , 2023
2023
Earlier work this paper cites.
S. Kim, S. Jung, B. Kim, M. Choi, J. Shin, and J. Lee, “Towards safe self-distillation of internet-scale text-to-image diffusion models,” in Proceedings of the International Conference on Machine Learning Workshops , 2023
2023
Earlier work this paper cites.
P. Schramowski, M. Brack, B. Deiseroth, and K. Kersting, “Safe latent diffusion: Mitigating inappropriate degeneration in diffusion models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Y. Qu, X. Shen, X. He, M. Backes, S. Zannettou, and Y. Zhang, “Unsafe diffusion: On the generation of unsafe images and hateful memes from text-to-image models,” in Proceedings of the ACM SIGSAC Conference on Computer and Communications Security , 2023
2023
Earlier work this paper cites.
OpenAI, “Openai usage policy,” 2023, accessed on 10-2023. [Online]. Available: https://openai.com/policies/usage-policies
2023
Earlier work this paper cites.
Meta, “Llama usage policy,” 2023, accessed on 10-2023. [Online]. Available: https://ai.meta.com/llama/use-policy
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
H. Liu, Y. Wu, S. Zhai, B. Yuan, and N. Zhang, “Riatig: Reliable and imperceptible adversarial text-to-image generation with natural prompts,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2023
2023
Earlier work this paper cites.
2024
Earlier work this paper cites.
2024
Cited alongside, same era.
Z. Kong, A. Goel, R. Badlani, W. Ping, R. Valle, and B. Catanzaro, “Audio flamingo: A novel audio language model with few-shot learning and dialogue abilities,” in Proceedings of the International Conference on Machine Learning , 2024
2024
Cited alongside, same era.
J. Xiao, A. Yao, Y. Li, and T.-S. Chua, “Can i trust your answer? visually grounded video question answering,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Cited alongside, same era.
X. Liu, P. P. Li, H. Huang, Z. Li, X. Cui, W. Deng, Z. He et al. , “Fka-owl: Advancing multimodal fake news detection through knowledge-augmented lvlms,” in Proceedings of the ACM International Conference on Multimedia , 2024
2024
Cited alongside, same era.
2024
Closest in time.
2024
Closest in time.
Y. Li, H. Guo, K. Zhou, W. X. Zhao, and J.-R. Wen, “Images are achilles’ heel of alignment: Exploiting visual vulnerabilities for jailbreaking multimodal large language models,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Closest in time.
X. Gu, X. Zheng, T. Pang, C. Du, Q. Liu, Y. Wang, J. Jiang, and M. Lin, “Agent smith: A single image can jailbreak one million multimodal llm agents exponentially fast,” in Proceedings of the International Conference on Machine Learning , 2024
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2024
Cited alongside, same era.
X. Cui, Z. Li, P. Li, H. Huang, X. Liu, and Z. He, “Instastyle: Inversion noise of a stylized image is secretly a style adviser,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Cited alongside, same era.
X. Cui, P. Li, Z. Li, X. Liu, Y. Zou, and Z. He, “Localize, understand, collaborate: Semantic-aware dragging via intention reasoner,” in Proceedings of the Advances in Neural Information Processing Systems , 2024
2024
Cited alongside, same era.
E. Shayegani, Y. Dong, and N. Abu-Ghazaleh, “Jailbreak in pieces: Compositional adversarial attacks on multi-modal language models,” in Proceedings of the International Conference on Learning Representations , 2024
2024
Cited alongside, same era.
Y. Yang, R. Gao, X. Wang, T.-Y. Ho, N. Xu, and Q. Xu, “Mma-diffusion: Multimodal attack on diffusion models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Cited alongside, same era.
Y. Miao, Y. Zhu, Y. Dong, L. Yu, J. Zhu, and X.-S. Gao, “T2vsafetybench: Evaluating the safety of text-to-video generative models,” in Proceedings of the Advances in Neural Information Processing Systems , 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
R. Wang, X. Ma, H. Zhou, C. Ji, G. Ye, and Y.-G. Jiang, “White-box multimodal jailbreaks against large vision-language models,” in Proceedings of the ACM International Conference on Multimedia , 2024
2024
Closest in time.
2024
Closest in time.
Z.-Y. Chin, C.-M. Jiang, C.-C. Huang, P.-Y. Chen, and W.-C. Chiu, “Prompting4debugging: Red-teaming text-to-image diffusion models by finding problematic prompts,” in Proceedings of the International Conference on Machine Learning , 2024
2024
Closest in time.
Y. Zhang, J. Jia, X. Chen, A. Chen, Y. Zhang, J. Liu, K. Ding, and S. Liu, “To generate or not? safety-driven unlearned diffusion models are still easy to generate unsafe images… for now,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Closest in time.
X. Li, Q. Shen, and K. Kawaguchi, “Va3: Virtually assured amplification attack on probabilistic copyright protection for text-to-image generative models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
R. Liu, A. Khakzar, J. Gu, Q. Chen, P. Torr, and F. Pizzati, “Latent guard: a safety framework for text-to-image generation,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Closest in time.
Y. Yang, R. Gao, X. Yang, J. Zhong, and Q. Xu, “Guardt2i: Defending text-to-image models from adversarial prompts,” in Proceedings of the Advances in Neural Information Processing Systems , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Z. Wu, H. Gao, Y. Wang, X. Zhang, and S. Wang, “Universal prompt optimizer for safe text-to-image generation,” in Proceedings of the North American Chapter of the Association for Computational Linguistics , 2024
2024
Closest in time.
Y. Zhang, X. Chen, J. Jia, Y. Zhang, C. Fan, J. Liu, M. Hong, K. Ding, and S. Liu, “Defensive unlearning with adversarial training for robust concept erasure in diffusion models,” in Proceedings of the Advances in Neural Information Processing Systems , 2024
2024
Closest in time.
H. Li, C. Shen, P. Torr, V. Tresp, and J. Gu, “Self-discovering interpretable diffusion latent directions for responsible text-to-image generation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Closest in time.
2024
Closest in time.
Y. Zong, O. Bohdal, T. Yu, Y. Yang, and T. Hospedales, “Safety fine-tuning at (almost) no cost: A baseline for vision large language models,” in Proceedings of the International Conference on Machine Learning , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
C. Fan, J. Liu, Y. Zhang, D. Wei, E. Wong, and S. Liu, “Salun: Empowering machine unlearning via gradient-based weight saliency in both image classification and generation,” in Proceedings of the International Conference on Learning Representations , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
X. Li, Y. Yang, J. Deng, C. Yan, Y. Chen, X. Ji, and W. Xu, “Safegen: Mitigating unsafe content generation in text-to-image models,” in Proceedings of the ACM Conference on Computer and Communications Security , 2024
2024
Closest in time.
R. Gandikota, H. Orgad, Y. Belinkov, J. Materzyńska, and D. Bau, “Unified concept editing in diffusion models,” in Proceedings of the IEEE Winter Conference on Applications of Computer Vision , 2024
2024
Closest in time.
S. Lu, Z. Wang, L. Li, Y. Liu, and A. W.-K. Kong, “Mace: Mass concept erasure in diffusion models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Closest in time.
C. Gong, K. Chen, Z. Wei, J. Chen, and Y.-G. Jiang, “Reliable and efficient concept erasure of text-to-image diffusion models,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Closest in time.
2024
Closest in time.
Y.-H. Park, S. Yun, J.-H. Kim, J. Kim, G. Jang, Y. Jeong, J. Jo, and G. Lee, “Direct unlearning optimization for robust and safe text-to-image models,” in Proceedings of the International Conference on Machine Learning Workshops , 2024
2024
Closest in time.
G. Zhang, K. Wang, X. Xu, Z. Wang, and H. Shi, “Forget-me-not: Learning to forget in text-to-image diffusion models,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2024
2024
Closest in time.
2024
Closest in time.
P. Wang, D. Zhang, L. Li, C. Tan, X. Wang, K. Ren, B. Jiang, and X. Qiu, “Inferaligner: Inference-time alignment for harmlessness through cross-model guidance,” in Proceedings of the Empirical Methods in Natural Language Processing , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
J. Gao, R. Pi, T. Han, H. Wu, L. Hong, L. Kong, X. Jiang, and Z. Li, “Coca: Regaining safety-awareness of multimodal large language models with constitutional calibration,” in Proceedings of the Conference on Language Modeling , 2024
2024
Closest in time.
2024
Closest in time.
R. Pi, T. Han, Y. Xie, R. Pan, Q. Lian, H. Dong, J. Zhang, and T. Zhang, “Mllm-protector: Ensuring mllm’s safety without hurting performance,” in Proceedings of the Empirical Methods in Natural Language Processing , 2024
2024
Closest in time.
Y. Gou, K. Chen, Z. Liu, L. Hong, H. Xu, Z. Li, D.-Y. Yeung, J. T. Kwok, and Y. Zhang, “Eyes closed, safety on: Protecting multimodal llms via image-to-text transformation,” in Proceedings of the European Conference on Computer Vision , 2024
2024
Closest in time.
S. Park, S. Moon, S. Park, and J. Kim, “Localization and manipulation of immoral visual cues for safe text-to-image generation,” in Proceedings of the IEEE Winter Conference on Applications of Computer Vision , 2024
2024
Closest in time.
L. Team, “Meta llama guard 2,” https://github.com/meta-llama/PurpleLlama/blob/main/Llama-Guard2/MODEL_CARD.md , 2024
2024
Closest in time.
2024
Closest in time.
M. Mazeika, L. Phan, X. Yin, A. Zou, Z. Wang, N. Mu, E. Sakhaee, N. Li, S. Basart, B. Li, D. A. Forsyth, and D. Hendrycks, “Harmbench: A standardized evaluation framework for automated red teaming and robust refusal,” in Proceedings of the International Conference on Machine Learning , 2024
2024
Closest in time.
2024
Closest in time.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenette,” 2024, https://github.com/fastai/imagenette
2024
Closest in time.
S. Christoph, K. Andreas, C. Theo, V. Richard, T. Benjamin, and R. Beaumont, “Laion-coco,” 2024, https://laion.ai/blog/laion-coco/
2024
Closest in time.