Fetching the paper…
Reading the bibliography…
We introduce ShieldGemma 2, a 4B parameter image content moderation model built on Gemma 3.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, et al · 2020
Earlier work this paper cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo · 2021
Earlier work this paper cites.
Zero-shot text-to-image generation
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, M. Chen, and I. Sutskever · 2021
Earlier work this paper cites.
Pali: A jointly-scaled multilingual language-image model
X. Chen, X. Wang, S. Changpinyo, A. Piergiovanni, P. Padlewski, D. Salz, S. Goodman, A. Grycner, B. Mustafa, L. Beyer, et al · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Earlier work this paper cites.
Photorealistic text-to-image diffusion models with deep language understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, et al · 2022
Earlier work this paper cites.
https://openai.com/research/gpt-4v-system-card , 2023
Gpt-4v · 2023
Earlier work this paper cites.
J. Achiam, S. Adler, S. Agarwal, L. Ahmad, I. Akkaya, F. L. Aleman, D. Almeida, J. Altenschmidt, S. Altman, S. Anadkat, et al · 2023
Earlier work this paper cites.
Gemini: a family of highly capable multimodal models
Gemini Team, R. Anil, S. Borgeaud, J.-B. Alayrac, J. Yu, R. Soricut, J. Schalkwyk, A. M. Dai, A. Hauth, K. Millican, et al · 2023
Earlier work this paper cites.
Figstep: Jailbreaking large vision-language models via typographic visual prompts
Y. Gong, D. Ran, J. Liu, C. Wang, T. Cong, A. Wang, S. Duan, and X. Wang · 2023
Earlier work this paper cites.
Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
J. Li, D. Li, S. Savarese, and S. Hoi · 2023
Cited alongside, same era.
Visual instruction tuning
H. Liu, C. Li, Q. Wu, and Y. J. Lee · 2023
Cited alongside, same era.
Safe latent diffusion: Mitigating inappropriate degeneration in diffusion models
P. Schramowski, M. Brack, B. Deiseroth, and K. Kersting · 2023
Cited alongside, same era.
S. Suri, F. Xiao, A. Sinha, S. C. Culatana, R. Krishnamoorthi, C. Zhu, and A. Shrivastava · 2023
Cited alongside, same era.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan · 2023
Cited alongside, same era.
Gemini 2 flash
Google · 2024
Later among the works it cites.
A survey on responsible generative ai: What to generate and what not
J. Gu · 2024
Later among the works it cites.
Llavaguard: Vlm-based safeguard for vision dataset curation and safety assessment
L. Helff, F. Friedrich, M. Brack, P. Schramowski, and K. Kersting · 2024
Later among the works it cites.
A. Hurst, A. Lerer, A. P. Goucher, A. Perelman, A. Ramesh, A. Clark, A. Ostrow, A. Welihinda, A. Hayes, A. Radford, et al · 2024
Later among the works it cites.
Self-discovering interpretable diffusion latent directions for responsible text-to-image generation
H. Li, C. Shen, P. Torr, V. Tresp, and J. Gu · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Baldridge, J. Bauer, M. Bhutani, N. Brichtova, A. Bunner, L. Castrejon, K. Chan, Y. Chen, S. Dieleman, Y. Du, et al · 2024
Cited alongside, same era.
Red teaming gpt-4v: Are gpt-4v safe against uni/multi-modal jailbreak attacks?
S. Chen, Z. Han, B. He, Z. Ding, W. Yu, P. Torr, V. Tresp, and J. Gu · 2024
Cited alongside, same era.
Uncovering vision modality threats in image-to-image tasks
H. Cheng, E. Xiao, J. Yang, J. Cao, Q. Zhang, J. Zhang, K. Xu, J. Gu, and R. Xu · 2024
Cited alongside, same era.
A. Dubey, A. Jauhri, A. Pandey, A. Kadian, A. Al-Dahle, A. Letman, A. Mathur, A. Schelten, A. Yang, A. Fan, et al · 2024
Cited alongside, same era.
Latent guard: a safety framework for text-to-image generation
R. Liu, A. Khakzar, J. Gu, Q. Chen, P. Torr, and F. Pizzati
Cited in the paper.
Multimodal pragmatic jailbreak on text-to-image models
T. Liu, Z. Lai, G. Zhang, P. Torr, V. Demberg, V. Tresp, and J. Gu
Cited in the paper.
Mm-safetybench: A benchmark for safety evaluation of multimodal large language models
X. Liu, Y. Zhu, J. Gu, Y. Lan, C. Yang, and Y. Qiao
Cited in the paper.
Y. Qu, X. Shen, Y. Wu, M. Backes, S. Zannettou, and Y. Zhang · 2024
Later among the works it cites.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
G. Team, P. Georgiev, V. I. Lei, R. Burnell, L. Bai, A. Gulati, G. Tanzer, D. Vincent, Z. Pan, S. Wang, et al · 2024
Later among the works it cites.
Shieldgemma: Generative ai content moderation based on gemma
W. Zeng, Y. Liu, R. Mullins, L. Peran, J. Fernandez, H. Harkous, K. Narasimhan, D. Proud, P. Kumar, B. Radharapu, et al · 2024
Later among the works it cites.
Orchestrating synthetic data with reasoning
T. R. Davidson, H. Harkous, B. Seguin, E. Bacis, and C. Ilharco · 2025
Closest in time.