Fetching the paper…
Reading the bibliography…
With the rapid advancement of Artificial Intelligence Generated Content (AIGC) technologies, synthetic images have become increasingly prevalent in everyday life, posing new challenges for authenticity assessment and detection.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
2013
Earlier work this paper cites.
S. Workman, R. Souvenir, and N. Jacobs, “Wide-area image geolocalization with aerial reference imagery,” in IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 1–9, acceptance rate: 30.3%
2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 618–626
2017
Earlier work this paper cites.
M. Sundararajan, A. Taly, and Q. Yan, “Axiomatic attribution for deep networks,” in International conference on machine learning . PMLR, 2017, pp. 3319–3328
2017
Earlier work this paper cites.
Y. Li, M.-C. Chang, and S. Lyu, “In ictu oculi: Exposing ai created fake videos by detecting eye blinking,” in WIFS , 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Rössler, D. Cozzolino, L. Verdoliva, C. Riess, J. Thies, and M. Nießner, “FaceForensics++: Learning to detect manipulated facial images,” in ICCV , 2019
2019
Earlier work this paper cites.
H. Lee, H.-E. Kim, and H. Nam, “Srm: A style-based recalibration module for convolutional neural networks,” in Proceedings of the IEEE/CVF International conference on computer vision , 2019, pp. 1854–1862
2019
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in neural information processing systems , vol. 33, pp. 6840–6851, 2020
2020
Earlier work this paper cites.
M. Barni, K. Kallas, E. Nowroozi, and B. Tondi, “Cnn detection of gan-generated face images based on cross-band co-occurrences analysis,” in 2020 IEEE international workshop on information forensics and security (WIFS) . IEEE, 2020, pp. 1–6
2020
Earlier work this paper cites.
S.-Y. Wang, O. Wang, R. Zhang, A. Owens, and A. A. Efros, “Cnn-generated images are surprisingly easy to spot… for now,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 8695–8704
2020
Earlier work this paper cites.
Z. Liu et al. , “Global texture enhancement for fake face detection in the wild,” in CVPR , 2020, pp. 8060–8069
2020
Earlier work this paper cites.
L. Li, J. Bao, T. Zhang, H. Yang, D. Chen, F. Wen, and B. Guo, “Face x-ray for more general face forgery detection,” in CVPR , 2020
2020
Earlier work this paper cites.
P. Dhariwal and A. Nichol, “Diffusion models beat gans on image synthesis,” Advances in neural information processing systems , vol. 34, pp. 8780–8794, 2021
2021
Earlier work this paper cites.
T. Zhao, X. Xu, M. Xu, H. Ding, Y. Xiong, and W. Xia, “Learning self-consistency for deepfake detection,” in Proceedings of the IEEE/CVF international conference on computer vision , 2021, pp. 15 023–15 033
2021
Earlier work this paper cites.
A. Haliassos, K. Vougioukas, S. Petridis, and M. Pantic, “Lips don’t lie: A generalisable and robust approach to face forgery detection,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 5039–5049
2021
Earlier work this paper cites.
S. Zhu, T. Yang, and C. Chen, “Vigor: Cross-view image geo-localization beyond one-to-one retrieval,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 3640–3649
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” in ICML , 2021
2021
Earlier work this paper cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, pp. 10 684–10 695
2022
Earlier work this paper cites.
M. Abdullah, A. Madain, and Y. Jararweh, “Chatgpt: Fundamentals, applications and social impacts,” in 2022 Ninth International Conference on Social Networks Analysis, Management and Security (SNAMS) . Ieee, 2022, pp. 1–8
2022
Earlier work this paper cites.
Y. Ju, S. Jia, L. Ke, H. Xue, K. Nagano, and S. Lyu, “Fusing global and local features for generalized ai-synthesized image detection,” in 2022 IEEE International Conference on Image Processing (ICIP) . IEEE, 2022, pp. 3465–3469
2022
Earlier work this paper cites.
M. Pawelec, “Deepfakes and democracy (theory): How synthetic audio-visual media for disinformation and hate speech threaten core democratic functions,” Digital society , vol. 1, no. 2, p. 19, 2022
2022
Earlier work this paper cites.
L. Zhang, A. Rao, and M. Agrawala, “Adding conditional control to text-to-image diffusion models,” in Proceedings of the IEEE/CVF international conference on computer vision , 2023, pp. 3836–3847
2023
Earlier work this paper cites.
H. H. Jiang, L. Brown, J. Cheng, M. Khan, A. Gupta, D. Workman, A. Hanna, J. Flowers, and T. Gebru, “Ai art and its impact on artists,” in Proceedings of the 2023 AAAI/ACM Conference on AI, Ethics, and Society , 2023, pp. 363–374
2023
Earlier work this paper cites.
G. Somepalli, V. Singla, M. Goldblum, J. Geiping, and T. Goldstein, “Diffusion art or digital forgery? investigating data replication in diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 6048–6058
2023
Cited alongside, same era.
D. Xu, S. Fan, and M. Kankanhalli, “Combating misinformation in the era of generative ai models,” in Proceedings of the 31st ACM International Conference on Multimedia , 2023, pp. 9291–9298
2023
Cited alongside, same era.
M. Mustak, J. Salminen, M. Mäntymäki, A. Rahman, and Y. K. Dwivedi, “Deepfakes: Deceptions, mitigations, and opportunities,” Journal of Business Research , vol. 154, p. 113368, 2023
2023
Cited alongside, same era.
U. Ojha, Y. Li, and Y. J. Lee, “Towards universal fake image detectors that generalize across generative models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 24 480–24 489
2023
Cited alongside, same era.
2024
Later among the works it cites.
M. Zhu, H. Chen, Q. Yan, X. Huang, G. Lin, W. Li, Z. Tu, H. Hu, J. Hu, and Y. Wang, “Genimage: A million-scale benchmark for detecting ai-generated image,” NeurIPS , vol. 36, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
J. Ye, Z. Lv, W. Li, J. Yu, H. Yang, H. Zhong, and C. He, “Cross-view image geo-localization with panorama-bev co-retrieval network,” in European Conference on Computer Vision . Springer, 2024, pp. 74–90
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Zhang, Z. Xu, C. Barnes, Y. Zhou, Q. Liu, H. Zhang, S. Amirghodsi, Z. Lin, E. Shechtman, and J. Shi, “Perceptual artifacts localization for image synthesis tasks,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 7579–7590
2023
Cited alongside, same era.
R. Shao, T. Wu, and Z. Liu, “Detecting and grounding multi-modal media manipulation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 6904–6913
2023
Cited alongside, same era.
B. Tu, X. Yang, W. He, J. Li, and A. Plaza, “Hyperspectral anomaly detection using reconstruction fusion of quaternion frequency domain analysis,” IEEE Transactions on Neural Networks and Learning Systems , 2023
2023
Cited alongside, same era.
F. Khalid, A. Javed, H. Ilyas, A. Irtaza et al. , “Dfgnn: An interpretable and generalized graph neural network for deepfakes detection,” Expert Systems with Applications , vol. 222, p. 119843, 2023
2023
Cited alongside, same era.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” 2023
2023
Cited alongside, same era.
H. Cheng, P. Zhang, S. Wu, J. Zhang, Q. Zhu, Z. Xie, J. Li, K. Ding, and L. Jin, “M6doc: A large-scale multi-format, multi-type, multi-layout, multi-language, multi-annotation category dataset for modern document layout analysis,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2023, pp. 15 138–15 147
2023
Cited alongside, same era.
R. Corvi, D. Cozzolino, G. Zingarini, G. Poggi, K. Nagano, and L. Verdoliva, “On the detection of synthetic images generated by diffusion models,” in ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2023, pp. 1–5
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Fang, Q. Sun, X. Wang, T. Huang, X. Wang, and Y. Cao, “Eva-02: A visual representation for neon genesis,” Image and Vision Computing , vol. 149, p. 105171, 2024
2024
Later among the works it cites.
Y. Zhang, A. Unell, X. Wang, D. Ghosh, Y. Su, L. Schmidt, and S. Yeung-Levy, “Why are visually-grounded language models bad at image classification?” Advances in Neural Information Processing Systems , vol. 37, pp. 51 727–51 753, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Tan, Y. Zhao, S. Wei, G. Gu, P. Liu, and Y. Wei, “Frequency-aware deepfake detection: Improving generalizability through frequency space domain learning,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 5, 2024, pp. 5052–5060
2024
Later among the works it cites.
H. Liu, Z. Tan, C. Tan, Y. Wei, J. Wang, and Y. Zhao, “Forgery-aware adaptive transformer for generalizable synthetic image detection,” in CVPR , 2024, pp. 10 770–10 780
2024
Later among the works it cites.
C. Tan, Y. Zhao, S. Wei, G. Gu, P. Liu, and Y. Wei, “Rethinking the up-sampling operations in cnn-based generative network for generalizable deepfake detection,” in CVPR , 2024, pp. 28 130–28 139
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Lin, W. Song, B. Li, Y. Li, J. Ni, H. Chen, and Q. Li, “Fake it till you make it: Curricular dynamic forgery augmentations towards general deepfake detection,” in European conference on computer vision . Springer, 2024, pp. 104–122
2024
Later among the works it cites.
2025
Closest in time.
J. Ye, B. Zhou, Z. Huang, J. Zhang, T. Bai, H. Kang, J. He, H. Lin, Z. Wang, T. Wu et al. , “Loki: A comprehensive synthetic data detection benchmark using large multimodal models,” in ICLR , 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
A. Yamada, T. Shiraishi, S. Nishino, T. Katsuoka, K. Taji, and I. Takeuchi, “Time series anomaly detection in the frequency domain with statistical reliability,” arXiv e-prints , pp. arXiv–2502, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
B. Zhou, H. Yang, D. Chen, J. Ye, T. Bai, J. Yu, S. Zhang, D. Lin, C. He, and W. Li, “Urbench: A comprehensive benchmark for evaluating large multimodal models in multi-view urban scenarios,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 39, no. 10, 2025, pp. 10 707–10 715
2025
Closest in time.
J. Ye, J. He, X. Zhang, Y. Lin, H. Lin, C. He, and W. Li, “Satellite image synthesis from street view with fine-grained spatial textual guidance: A novel framework,” IEEE Geoscience and Remote Sensing Magazine , 2025
2025
Closest in time.