Fetching the paper…
Reading the bibliography…
Deep generative models can create remarkably photorealistic fake images while raising concerns about misinformation and copyright infringement, known as deepfake threats.
Detecting GAN generated fake images using co-occurrence matrices
Nataraj, L.; Mohammed, T. M.; Chandrasekaran, S.; Flenner, A.; Bappy, J. H.; Roy-Chowdhury, A. K.; and Manjunath, B. 2019 · 1903
Earlier work this paper cites.
The Deepfake Detection Challenge (DFDC) Preview Dataset
Dolhansky, B.; Howes, R.; Pflaum, B.; Baram, N.; and Ferrer, C. C. 2019 · 1910
Earlier work this paper cites.
Faceshifter: Towards high fidelity and occlusion aware face swapping
Li, L.; Bao, J.; Yang, H.; Chen, D.; and Wen, F. 2019 · 1912
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
Witches’ brew: Industrial scale data poisoning via gradient matching
Geiping, J.; Fowl, L.; Huang, W. R.; Czaja, W.; Taylor, G.; Moeller, M.; and Goldstein, T. 2020 · 2009
Earlier work this paper cites.
Torchattacks: A pytorch repository for adversarial attacks
Kim, H. 2020 · 2010
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2010
Earlier work this paper cites.
Generative Adversarial Nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y.; Maire, M.; Belongie, S.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Young, P.; Lai, A.; Hodosh, M.; and Hockenmaier, J. 2014 · 2014
Earlier work this paper cites.
Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop
Yu, F.; Seff, A.; Zhang, Y.; Song, S.; Funkhouser, T.; and Xiao, J. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Face2face: Real-time face capture and reenactment of rgb videos
Thies, J.; Zollhofer, M.; Stamminger, M.; Theobalt, C.; and Nießner, M. 2016 · 2016
Earlier work this paper cites.
Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Goyal, Y.; Khot, T.; Summers-Stay, D.; Batra, D.; and Parikh, D. 2017 · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Loshchilov, I.; and Hutter, F. 2017 · 2017
Earlier work this paper cites.
Large scale GAN training for high fidelity natural image synthesis
Brock, A.; Donahue, J.; and Simonyan, K. 2018 · 2018
Earlier work this paper cites.
Vizwiz grand challenge: Answering visual questions from blind people
Gurari, D.; Li, Q.; Stangl, A. J.; Guo, A.; Lin, C.; Grauman, K.; Luo, J.; and Bigham, J. P. 2018 · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Karras, T.; Laine, S.; and Aila, T. 2019 · 2019
Earlier work this paper cites.
Ok-vqa: A visual question answering benchmark requiring external knowledge
Marino, K.; Rastegari, M.; Farhadi, A.; and Mottaghi, R. 2019 · 2019
Earlier work this paper cites.
Ocr-vqa: Visual question answering by reading text in images
Mishra, A.; Shekhar, S.; Singh, A. K.; and Chakraborty, A. 2019 · 2019
Earlier work this paper cites.
Faceforensics++: Learning to detect manipulated facial images
Rossler, A.; Cozzolino, D.; Verdoliva, L.; Riess, C.; Thies, J.; and Nießner, M. 2019 · 2019
Earlier work this paper cites.
Towards vqa models that can read
Singh, A.; Natarajan, V.; Shah, M.; Jiang, Y.; Chen, X.; Batra, D.; Parikh, D.; and Rohrbach, M. 2019 · 2019
Earlier work this paper cites.
Deferred neural rendering: Image synthesis using neural textures
Thies, J.; Zollhöfer, M.; and Nießner, M. 2019 · 2019
Earlier work this paper cites.
Attributing fake images to gans: Learning and analyzing gan fingerprints
Yu, N.; Davis, L. S.; and Fritz, M. 2019 · 2019
Earlier work this paper cites.
Self-attention generative adversarial networks
Zhang, H.; Goodfellow, I.; Metaxas, D.; and Odena, A. 2019 · 2019
Earlier work this paper cites.
Detecting and simulating artifacts in gan fake images
Zhang, X.; Karaman, S.; and Chang, S.-F. 2019 · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Cited alongside, same era.
Deeperforensics-1.0: A large-scale dataset for real-world face forgery detection
Jiang, L.; Li, R.; Wu, W.; Qian, C.; and Loy, C. C. 2020 · 2020
Cited alongside, same era.
Analyzing and improving the image quality of stylegan
Karras, T.; Laine, S.; Aittala, M.; Hellsten, J.; Lehtinen, J.; and Aila, T. 2020 · 2020
Cited alongside, same era.
Celeb-df: A large-scale challenging dataset for deepfake forensics
Li, Y.; Yang, X.; Sun, P.; Qi, H.; and Lyu, S. 2020 · 2020
Cited alongside, same era.
CNN-generated images are surprisingly easy to spot… for now
Wang, S.-Y.; Wang, O.; Zhang, R.; Owens, A.; and Efros, A. A. 2020 · 2020
Cited alongside, same era.
Danbooru: A large-scale crowd sourced and tagged anime illustration dataset
2021 · 2021
Scienceqa: A novel resource for question answering on scholarly articles
Saikh, T.; Ghosal, T.; Mittal, A.; Ekbal, A.; and Bhattacharyya, P. 2022 · 2022
Later among the works it cites.
StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets
Sauer, A.; Schwarz, K.; and Geiger, A. 2022 · 2022
Later among the works it cites.
A-okvqa: A benchmark for visual question answering using world knowledge
Schwenk, D.; Khandelwal, A.; Clark, C.; Marino, K.; and Mottaghi, R. 2022 · 2022
Later among the works it cites.
De-fake: Detection and attribution of fake images generated by text-to-image diffusion models
Sha, Z.; Li, Z.; Yu, N.; and Zhang, Y. 2022 · 2022
Later among the works it cites.
Resolution-robust large mask inpainting with fourier convolutions
Suvorov, R.; Logacheva, E.; Mashikhin, A.; Remizova, A.; Ashukha, A.; Silvestrov, A.; Kong, N.; Goka, H.; Park, K.; and Lempitsky, V. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning continuous image representation with local implicit image function
Chen, Y.; Liu, S.; and Wang, X. 2021 · 2021
Cited alongside, same era.
Diffusion Models Beat GANs on Image Synthesis
Dhariwal, P.; and Nichol, A. 2021 · 2021
Cited alongside, same era.
Fighting deepfakes by detecting gan dct anomalies
Giudice, O.; Guarnera, L.; and Battiato, S. 2021 · 2021
Cited alongside, same era.
Beyond the spectrum: Detecting deepfakes via re-synthesis
He, Y.; Yu, N.; Keuper, M.; and Fritz, M. 2021 · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Hu, E. J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W. 2021 · 2021
Cited alongside, same era.
Alias-Free Generative Adversarial Networks
Karras, T.; Aittala, M.; Laine, S.; Härkönen, E.; Hellsten, J.; Lehtinen, J.; and Aila, T. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
Tempera: Test-time prompt editing via reinforcement learning
Zhang, T.; Wang, X.; Zhou, D.; Schuurmans, D.; and Gonzalez, J. E. 2022 · 2022
Later among the works it cites.
InstructZero: Efficient Instruction Optimization for Black-Box Large Language Models
Chen, L.; Chen, J.; Goldstein, T.; Huang, H.; and Zhou, T. 2023 · 2023
Closest in time.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; et al. 2023 · 2023
Closest in time.
InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Dai, W.; Li, J.; Li, D.; Tiong, A. M. H.; Zhao, J.; Wang, W.; Li, B.; Fung, P.; and Hoi, S. 2023 · 2023
Closest in time.
Guarnera, L.; Giudice, O.; and Battiato, S. 2023 · 2023
Closest in time.
Quality-Agnostic Deepfake Detection with Intra-model Collaborative Learning
Le, B. M.; and Woo, S. S. 2023 · 2023
Closest in time.
Li, J.; Li, D.; Savarese, S.; and Hoi, S. 2023 · 2023
Closest in time.
Detecting Images Generated by Deep Diffusion Models using their Local Intrinsic Dimensionality
Lorenz, P.; Durall, R. L.; and Keuper, J. 2023 · 2023
Closest in time.
Exposing the Fake: Effective Diffusion-Generated Images Detection
Ma, R.; Duan, J.; Kong, F.; Shi, X.; and Xu, K. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
SDXL: improving latent diffusion models for high-resolution image synthesis
Podell, D.; English, Z.; Lacey, K.; Blattmann, A.; Dockhorn, T.; Müller, J.; Penna, J.; and Rombach, R. 2023 · 2023
Closest in time.
DeepFloyd IF
StabilityAI. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; Azhar, F.; et al. 2023 · 2023
Closest in time.
Hard prompts made easy: Gradient-based discrete optimization for prompt tuning and discovery
Wen, Y.; Jain, N.; Kirchenbauer, J.; Goldblum, M.; Geiping, J.; and Goldstein, T. 2023 · 2023
Closest in time.
Generalizable Synthetic Image Detection via Language-guided Contrastive Learning
Wu, H.; Zhou, J.; and Zhang, S. 2023 · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Zhang, L.; and Agrawala, M. 2023 · 2023
Closest in time.
Minigpt-4: Enhancing vision-language understanding with advanced large language models
Zhu, D.; Chen, J.; Shen, X.; Li, X.; and Elhoseiny, M. 2023 · 2023
Closest in time.
Universal and transferable adversarial attacks on aligned language models
Zou, A.; Wang, Z.; Kolter, J. Z.; and Fredrikson, M. 2023 · 2023
Closest in time.
Playground v2. 5: Three insights towards enhancing aesthetic quality in text-to-image generation
Li, D.; Kamko, A.; Akhgari, E.; Sabet, A.; Xu, L.; and Doshi, S. 2024 · 2024
Closest in time.
Forgery-aware adaptive transformer for generalizable synthetic image detection
Liu, H.; Tan, Z.; Tan, C.; Wei, Y.; Wang, J.; and Zhao, Y. 2024 · 2024
Closest in time.