Fetching the paper…
Reading the bibliography…
The rapid development of generative AI is a double-edged sword, which not only facilitates content creation but also makes image manipulation easier and more difficult to detect.
Columbia image splicing detection evaluation dataset
Tian-Tsong Ng, Jessie Hsu, and Shih-Fu Chang · 2009
Earlier work this paper cites.
Exposing digital image forgeries by illumination color classification
Tiago José De Carvalho, Christian Riess, Elli Angelopoulou, Helio Pedrini, and Anderson de Rezende Rocha · 2013
Earlier work this paper cites.
Casia image tampering detection evaluation database
Jing Dong, Wei Wang, and Tieniu Tan · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Multi-scale analysis strategies in prnu-based tampering localization
Paweł Korus and Jiwu Huang · 2016
Earlier work this paper cites.
Coverage—a novel database for copy-move forgery detection
Bihan Wen, Ye Zhu, Ramanathan Subramanian, Tian-Tsong Ng, Xuanjing Shen, and Stefan Winkler · 2016
Earlier work this paper cites.
Generalised dice overlap as a deep learning loss function for highly unbalanced segmentations
Carole H Sudre, Wenqi Li, Tom Vercauteren, Sebastien Ourselin, and M Jorge Cardoso · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin · 2018
Earlier work this paper cites.
Fast and effective image copy-move forgery detection via hierarchical feature point matching
Yuanman Li and Jiantao Zhou · 2018
Earlier work this paper cites.
Learning a convolutional neural network for image compact-resolution
Yue Li, Dong Liu, Houqiang Li, Li Li, Zhu Li, and Feng Wu · 2018
Earlier work this paper cites.
Image splicing localization using a multi-task fully convolutional network (mfcn)
Ronald Salloum, Yuzhuo Ren, and C-C Jay Kuo · 2018
Earlier work this paper cites.
A deep learning approach to patch-based image inpainting forensics
Xinshan Zhu, Yongjun Qian, Xianfeng Zhao, Biao Sun, and Ya Sun · 2018
Earlier work this paper cites.
The point where reality meets fantasy: Mixed adversarial generators for image splice detection
Vladimir V Kniaz, Vladimir Knyaz, and Fabio Remondino · 2019
Earlier work this paper cites.
Localization of deep inpainting using high-pass fully convolutional network
Haodong Li and Jiwu Huang · 2019
Earlier work this paper cites.
Mantra-net: Manipulation tracing network for detection and localization of image forgeries with anomalous features
Yue Wu, Wael AbdAlmageed, and Premkumar Natarajan · 2019
Earlier work this paper cites.
On the detection of digital face manipulation
Hao Dang, Feng Liu, Joel Stehouwer, Xiaoming Liu, and Anil K Jain · 2020
Earlier work this paper cites.
Span: Spatial pyramid attention network for image manipulation localization
Xuefeng Hu, Zhihan Zhang, Zhenye Jiang, Syomantak Chaudhuri, Zhenheng Yang, and Ram Nevatia · 2020
Earlier work this paper cites.
Doa-gan: Dual-order attentive generative adversarial network for image copy-move forgery detection and localization
Ashraful Islam, Chengjiang Long, Arslan Basharat, and Anthony Hoogs · 2020
Earlier work this paper cites.
Imd2020: A large-scale annotated dataset tailored for detecting manipulated images
Adam Novozamsky, Babak Mahdian, and Stanislav Saic · 2020
Earlier work this paper cites.
Image manipulation detection by multi-view multi-scale supervision
Xinru Chen, Chengbo Dong, Jiaqi Ji, Juan Cao, and Xirong Li · 2021
Earlier work this paper cites.
Cat-net: Compression artifact tracing network for detection and localization of image splicing
Myung-Joon Kwon, In-Jae Yu, Seung-Hun Nam, and Heung-Kyu Lee · 2021
Cited alongside, same era.
Deepfake detection based on discrepancies between faces and their context
Yuval Nirkin, Lior Wolf, Yosi Keller, and Tal Hassner · 2021
Cited alongside, same era.
From image to imuge: Immunized image generation
Qichao Ying, Zhenxing Qian, Hang Zhou, Haisheng Xu, Xinpeng Zhang, and Siyi Li · 2021
Cited alongside, same era.
End-to-end reconstruction-classification learning for face forgery detection
Junyi Cao, Chao Ma, Taiping Yao, Shen Chen, Shouhong Ding, and Xiaokang Yang · 2022
Cited alongside, same era.
Mvss-net: Multi-view multi-scale supervised networks for image manipulation detection
Chengbo Dong, Xinru Chen, Ruohan Hu, Juan Cao, and Xirong Li · 2022
Cited alongside, same era.
LoRA: Low-rank adaptation of large language models
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Iml-vit: Image manipulation localization by vision transformer
Xiaochen Ma, Bo Du, Xianggen Liu, Ahmed Y Al Hammadi, and Jizhe Zhou · 2023
Later among the works it cites.
Dragondiffusion: Enabling drag-style manipulation on diffusion models
Chong Mou, Xintao Wang, Jiechong Song, Ying Shan, and Jian Zhang · 2023
Later among the works it cites.
Gpt-4 technical report. arxiv 2303.08774
R OpenAI · 2023
Later among the works it cites.
Sdxl: improving latent diffusion models for high-resolution image synthesis
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Edward J Hu, belong shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2022
Cited alongside, same era.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi · 2022
Cited alongside, same era.
Pscc-net: Progressive spatio-channel correlation network for image manipulation detection and localization
Xiaohong Liu, Yaojie Liu, Jun Chen, and Xiaoming Liu · 2022
Cited alongside, same era.
Repaint: Inpainting using denoising diffusion probabilistic models
Andreas Lugmayr, Martin Danelljan, Andres Romero, Fisher Yu, Radu Timofte, and Luc Van Gool · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Cited alongside, same era.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, Stephen Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Arun Raja, Manan Dey, et al · 2022
Cited alongside, same era.
Detecting and recovering sequential deepfake manipulation
Rui Shao, Tianxing Wu, and Ziwei Liu · 2022
Cited alongside, same era.
Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas Müller, Joe Penna, and Robin Rombach · 2023
Later among the works it cites.
Cogvlm: Visual expert for pretrained language models
Weihan Wang, Qingsong Lv, Wenmeng Yu, Wenyi Hong, Ji Qi, Yan Wang, Junhui Ji, Zhuoyi Yang, Lei Zhao, Xixuan Song, et al · 2023
Later among the works it cites.
Learning to immunize images for tamper localization and self-recovery
Qichao Ying, Hang Zhou, Zhenxing Qian, Sheng Li, and Xinpeng Zhang · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala · 2023
Later among the works it cites.
Exposing the deception: Uncovering more forgery clues for deepfake detection
Zhongjie Ba, Qingyu Liu, Zhenguang Liu, Shuang Wu, Feng Lin, Li Lu, and Kui Ren · 2024
Closest in time.
Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, Qinglong Zhang, Xizhou Zhu, Lewei Lu, et al · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al · 2024
Closest in time.
Ffaa: Multimodal large language model based explainable open-world face forgery analysis assistant
Zhengchao Huang, Bin Xia, Zicheng Lin, Zhun Mou, and Wenming Yang · 2024
Closest in time.
Lisa: Reasoning segmentation via large language model
Xin Lai, Zhuotao Tian, Yukang Chen, Yanwei Li, Yuhui Yuan, Shu Liu, and Jiaya Jia · 2024
Closest in time.
Protect-your-ip: Scalable source-tracing and attribution against personalized generation
Runyi Li, Xuanyu Zhang, Zhipei Xu, Yongbing Zhang, and Jian Zhang · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2024
Closest in time.
Flickr-faces-hq dataset (ffhq)
NVIDIA Corporation · 2024
Closest in time.
Glamm: Pixel grounding large multimodal model
Hanoona Rasheed, Muhammad Maaz, Sahal Shaji, Abdelrahman Shaker, Salman Khan, Hisham Cholakkal, Rao M Anwer, Eric Xing, Ming-Hsuan Yang, and Fahad S Khan · 2024
Closest in time.
Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution
Peng Wang, Shuai Bai, Sinan Tan, Shijie Wang, Zhihao Fan, Jinze Bai, Keqin Chen, Xuejing Liu, Jialin Wang, Wenbin Ge, et al · 2024
Closest in time.
Research about the ability of llm in the tamper-detection area
Xinyu Yang and Jizhe Zhou · 2024
Closest in time.