Fetching the paper…
Reading the bibliography…
The ability to distinguish whether an image is generated by artificial intelligence (AI) is a crucial ingredient in human intelligence, usually accompanied by a complex and dialectical forensic and reasoning process.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: A method for automatic evaluation of machine translation,” in ACL , 2002, pp. 311–318
2002
Earlier work this paper cites.
C.-Y. Lin, “Rouge: A package for automatic evaluation of summaries,” in ACL , 2004, pp. 74–81
2004
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: A large-scale hierarchical image database,” in CVPR . IEEE, 2009, pp. 248–255
2009
Earlier work this paper cites.
2015
Earlier work this paper cites.
E. Agustsson and R. Timofte, “Ntire 2017 challenge on single image super-resolution: Dataset and study,” in CVPRW . IEEE, 2017, pp. 126–135
2017
Earlier work this paper cites.
T. Karras, T. Aila, S. Laine, and J. Lehtinen, “Progressive growing of GANs for improved quality, stability, and variation,” in ICLR , 2018
2018
Earlier work this paper cites.
A. Talmor, J. Herzig, N. Lourie, and J. Berant, “CommonsenseQA: A question answering challenge targeting commonsense knowledge,” in ACL , 2018, pp. 4149–4158
2018
Earlier work this paper cites.
T. Karras, S. Laine, and T. Aila, “A style-based generator architecture for generative adversarial networks,” in CVRP . IEEE, 2019, pp. 4401–4410
2019
Earlier work this paper cites.
N. Reimers and I. Gurevych, “Sentence-BERT: Sentence embeddings using siamese bert-networks,” in EMNLP , 2019, pp. 3982–3992
2019
Earlier work this paper cites.
M. Barni, K. Kallas, E. Nowroozi, and B. Tondi, “CNN detection of GAN-generated face images based on cross-band co-occurrences analysis,” in WIFS . IEEE, 2020, pp. 1–6
2020
Earlier work this paper cites.
J. Frank, T. Eisenhofer, L. Schönherr, A. Fischer, D. Kolossa, and T. Holz, “Leveraging frequency analysis for deep fake image recognition,” in ICML . PMLR, 2020, pp. 3247–3258
2020
Earlier work this paper cites.
Z. Liu, X. Qi, and P. H. Torr, “Global texture enhancement for fake face detection in the wild,” in CVPR . IEEE, 2020, pp. 8060–8069
2020
Earlier work this paper cites.
L. Chai, D. Bau, S.-N. Lim, and P. Isola, “What makes fake images detectable? Understanding properties that generalize,” in ECCV . Springer, 2020, pp. 103–120
2020
Earlier work this paper cites.
S.-Y. Wang, O. Wang, R. Zhang, A. Owens, and A. A. Efros, “CNN-generated images are surprisingly easy to spot… for now,” in CVPR . IEEE, 2020, pp. 8695–8704
2020
Earlier work this paper cites.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly et al. , “An image is worth 16x16 words: Transformers for image recognition at scale,” ICLR , 2020
2020
Earlier work this paper cites.
R. Wang, F. Juefei-Xu, L. Ma, X. Xie, Y. Huang, J. Wang, and Y. Liu, “Fakespotter: A simple yet robust baseline for spotting ai-synthesized fake faces,” IJCAI , pp. 3444–3451, 2020
2020
Earlier work this paper cites.
D. Gragnaniello, D. Cozzolino, F. Marra, G. Poggi, and L. Verdoliva, “Are GAN generated images easy to detect? A critical analysis of the state-of-the-art,” in ICME . IEEE, 2021, pp. 1–6
2021
Earlier work this paper cites.
Y. He, B. Gan, S. Chen, Y. Zhou, G. Yin, L. Song, L. Sheng, J. Shao, and Z. Liu, “Forgerynet: A versatile benchmark for comprehensive forgery analysis,” in CVPR . IEEE, 2021, pp. 4360–4369
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Y. Ju, S. Jia, L. Ke, H. Xue, K. Nagano, and S. Lyu, “Fusing global and local features for generalized AI-synthesized image detection,” in ICIP . IEEE, 2022, pp. 3465–3469
2022
Earlier work this paper cites.
B. Liu, F. Yang, X. Bi, B. Xiao, W. Li, and X. Gao, “Detecting generated images by real images,” in ECCV . Springer, 2022, pp. 95–110
2022
Earlier work this paper cites.
S. Dong, J. Wang, J. Liang, H. Fan, and R. Ji, “Explaining deepfake detection by analysing image matching,” in ECCV . Springer, 2022, pp. 18–35
2022
Earlier work this paper cites.
M. Wolter, F. Blanke, R. Heese, and J. Garcke, “Wavelet-packets for deepfake image analysis and detection,” Mach. Learn. , vol. 111, pp. 4295–4327, 2022
2022
Earlier work this paper cites.
L. Verdoliva, D. Cozzolino, and K. Nagano, “2022 ieee image and video processing cup synthetic image detection,” IEEE , 2022
2022
Earlier work this paper cites.
M. Ding, W. Zheng, W. Hong, and J. Tang, “Cogview2: Faster and better text-to-image generation via hierarchical transformers,” in NeurIPS , 2022, pp. 16 890–16 902
2022
Earlier work this paper cites.
S. Gu, D. Chen, J. Bao, F. Wen, B. Zhang, D. Chen, L. Yuan, and B. Guo, “Vector quantized diffusion model for text-to-image synthesis,” in CVPR . IEEE, 2022, pp. 10 696–10 706
2022
Earlier work this paper cites.
A. Nichol, P. Dhariwal, A. Ramesh, P. Shyam, P. Mishkin, B. McGrew, I. Sutskever, and M. Chen, “Glide: Towards photorealistic image generation and editing with text-guided diffusion models,” in ICML . PMLR, 2022, pp. 16 784–16 804
2022
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in CVPR . IEEE, 2022, pp. 10 684–10 695
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
P. Lu, S. Mishra, T. Xia, L. Qiu, K.-W. Chang, S.-C. Zhu, O. Tafjord, P. Clark, and A. Kalyan, “Learn to explain: Multimodal reasoning via thought chains for science question answering,” in NeurIPS , 2022, pp. 2507–2521
2022
Earlier work this paper cites.
H. Farid, “Lighting (in) consistency of paint by text,” arXiv preprint arXiv:2207.13744 , 2022
2022
Earlier work this paper cites.
——, “Perspective (in) consistency of paint by text,” arXiv preprint arXiv:2206.14617 , 2022
2022
Earlier work this paper cites.
O. Rubin, J. Herzig, and J. Berant, “Learning to retrieve prompts for in-context learning,” in ACL , 2022, pp. 2655–2671
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, b. ichter, F. Xia, E. Chi, Q. V. Le, and D. Zhou, “Chain-of-thought prompting elicits reasoning in large language models,” in NeurIPS , 2022, pp. 24 824–24 837
2022
Earlier work this paper cites.
X. Wu, L. Xiao, Y. Sun, J. Zhang, T. Ma, and L. He, “A survey of human-in-the-loop for machine learning,” Future Gener. Comp. Sy. , vol. 135, pp. 364–381, 2022
2022
Cited alongside, same era.
Z. Du, Y. Qian, X. Liu, M. Ding, J. Qiu, Z. Yang, and J. Tang, “GLM: General language model pretraining with autoregressive blank infilling,” in ACL , 2022, pp. 320–335
2022
Cited alongside, same era.
Z. Zhang, A. Zhang, M. Li, and A. Smola, “Automatic chain of thought prompting in large language models,” in EMNLP , 2022, pp. 12 113–12 139
2022
Cited alongside, same era.
J. Betker, G. Goh, L. Jing, T. Brooks, J. Wang, L. Li, L. Ouyang, J. Zhuang, J. Lee, Y. Guo et al. , “Improving image generation with better captions,” Computer Science , vol. 2, no. 3, p. 8, 2023
2023
Cited alongside, same era.
C. Tan, Y. Zhao, S. Wei, G. Gu, and Y. Wei, “Learning on gradients: Generalized artifacts representation for gan-generated images detection,” in CVPR . IEEE, 2023, pp. 12 105–12 114
H. Laurençon, L. Saulnier, L. Tronchon, S. Bekman, A. Singh, A. Lozhkov, T. Wang, S. Karamcheti, A. M. Rush, D. Kiela, M. Cord, and V. Sanh, “OBELICS: An open web-scale filtered dataset of interleaved image-text documents,” in NeurIPS , 2023
2023
Later among the works it cites.
Z. Peng, W. Wang, L. Dong, Y. Hao, S. Huang, S. Ma, and F. Wei, “Kosmos-2: Grounding multimodal large language models to the world,” arXiv preprint arXiv: 2306.14824 , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
A. Aghasanli, D. Kangin, and P. Angelov, “Interpretable-through-prototypes deepfake detection for diffusion models,” in ICCVW . IEEE, 2023, pp. 467–474
2023
Cited alongside, same era.
L. Zhang, Z. Xu, C. Barnes, Y. Zhou, Q. Liu, H. Zhang, S. Amirghodsi, Z. Lin, E. Shechtman, and J. Shi, “Perceptual artifacts localization for image synthesis tasks,” in CVPR . IEEE, 2023, pp. 7579–7590
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Z. Wang, J. Bao, W. Zhou, W. Wang, H. Hu, H. Chen, and H. Li, “Dire for diffusion-generated image detection,” in ICCV . IEEE, 2023, pp. 22 388–22 398
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Hou, Q. Guo, Y. Huang, X. Xie, L. Ma, and J. Zhao, “Evading deepfake detectors via adversarial statistical consistency,” in CVPR . IEEE, 2023, pp. 12 271–12 280
2023
Cited alongside, same era.
U. Ojha, Y. Li, and Y. J. Lee, “Towards universal fake image detectors that generalize across generative models,” in CVPR . IEEE, 2023, pp. 24 480–24 489
2023
Cited alongside, same era.
2023
Later among the works it cites.
“Midjourney,” 2024. [Online]. Available: https://www.midjourney.com/home
2024
Closest in time.
D. Tariang, R. Corvi, D. Cozzolino, G. Poggi, K. Nagano, and L. Verdoliva, “Synthetic image verification in the era of generative artificial intelligence: What works and what isn’t there yet,” IEEE Security & Privacy , 2024
2024
Closest in time.
J. J. Bird and A. Lotfi, “Cifake: Image classification and explainable identification of ai-generated synthetic images,” IEEE Access , 2024
2024
Closest in time.
Y. Zhang, B. Colman, A. Shahriyari, and G. Bharaj, “Common sense reasoning for deep fake detection,” in ECCV . Springer, 2024, pp. 1–1
2024
Closest in time.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” in NeurIPS , vol. 36, 2024
2024
Closest in time.
Y. Ge, W. Hua, K. Mei, J. Tan, S. Xu, Z. Li, Y. Zhang et al. , “OpenAGI: When LLM meets domain experts,” in NeurIPS , vol. 36, 2024
2024
Closest in time.
H. Zhu, X. Sui, B. Chen, X. Liu, P. Chen, Y. Fang, and S. Wang, “2AFC prompting of large multimodal models for image quality assessment,” IEEE Trans. Circuits Syst. Video Technol. , pp. 1–1, 2024
2024
Closest in time.
H. Zhao, H. Chen, F. Yang, N. Liu, H. Deng, H. Cai, S. Wang, D. Yin, and M. Du, “Explainability for large language models: A survey,” ACM Trans. Intell. Syst. Technol. , vol. 15, no. 2, pp. 1–38, 2024
2024
Closest in time.
2024
Closest in time.
Y. Ju, S. Jia, J. Cai, H. Guan, and S. Lyu, “GLFF: Global and local feature fusion for ai-synthesized image detection,” IEEE Trans. Multimedia , vol. 26, pp. 4073–4085, 2024
2024
Closest in time.
A. Sarkar, H. Mai, A. Mahapatra, S. Lazebnik, D. A. Forsyth, and A. Bhattad, “Shadows don’t lie and lines can’t bend! generative models don’t know projective geometry… for now,” in CVPR . IEEE, 2024, pp. 28 140–28 149
2024
Closest in time.
Z. Lu, D. Huang, L. Bai, J. Qu, C. Wu, X. Liu, and W. Ouyang, “Seeing is not always believing: benchmarking human and model perception of ai-generated images,” NeurIPS , vol. 36, 2024
2024
Closest in time.
M. Zhu, H. Chen, Q. Yan, X. Huang, G. Lin, W. Li, Z. Tu, H. Hu, J. Hu, and Y. Wang, “Genimage: A million-scale benchmark for detecting ai-generated image,” in NeurIPS , 2024
2024
Closest in time.
2024
Closest in time.
“Dalle3 reddit dataset,” 2024. [Online]. Available: https://huggingface.co/datasets/ProGamerGov/dalle-3-reddit-dataset
2024
Closest in time.
“midjourney-v5 prompt dataset,” 2024. [Online]. Available: https://huggingface.co/datasets/tarungupta83/MidJourney_v5_Prompt_dataset
2024
Closest in time.
D. Zhu, J. Chen, X. Shen, X. Li, and M. Elhoseiny, “Minigpt-4: Enhancing vision-language understanding with advanced large language models,” in ICLR , 2024
2024
Closest in time.
2024
Closest in time.
Anthrop, “Introducing the next generation of claude,” https://www.anthropic.com/news/claude-3-family , 2024.03
2024
Closest in time.
B. Li, Y. Ge, Y. Ge, G. Wang, R. Wang, R. Zhang, and Y. Shan, “Seed-bench: Benchmarking multimodal large language models,” in CVPR . IEEE, 2024, pp. 13 299–13 308
2024
Closest in time.
2024
Closest in time.
Y. Huang, Q. Yuan, X. Sheng, Z. Yang, H. Wu, P. Chen, Y. Yang, L. Li, and W. Lin, “AesBench: An expert benchmark for multimodal large language models on image aesthetics perception,” in ACM MM . ACM, 2024
2024
Closest in time.
H. Park, G. Kim, D. Lee, and H. K. Kim, “Can you spot the ai-generated images? distinguishing fake images using signal detection theory,” in HCI International . Springer, 2024, pp. 299–313
2024
Closest in time.
H. Wu, Z. Zhang, E. Zhang, C. Chen, L. Liao, A. Wang, K. Xu, C. Li, J. Hou, G. Zhai et al. , “Q-Instruct: Improving low-level visual abilities for multi-modality foundation models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 25 490–25 500
2024
Closest in time.
2024
Closest in time.
L. Zheng, W.-L. Chiang, Y. Sheng, S. Zhuang, Z. Wu, Y. Zhuang, Z. Lin, Z. Li, D. Li, E. Xing et al. , “Judging LLM-as-a-judge with MT-bench and chatbot arena,” in NeurIPS , vol. 36, 2024
2024
Closest in time.
Z. Zhang, A. Zhang, M. Li, H. Zhao, G. Karypis, and A. Smola, “Multimodal chain-of-thought reasoning in language models,” Trans. Mach. Learn. Res. , vol. 2024, 2024
2024
Closest in time.
H. Wu, Z. Zhang, E. Zhang, C. Chen, L. Liao, A. Wang, C. Li, W. Sun, Q. Yan, G. Zhai et al. , “Q-bench: A benchmark for general-purpose foundation models on low-level vision,” in ICLR , 2024
2024
Closest in time.