Fetching the paper…
Reading the bibliography…
The fast evolution and widespread of deepfake techniques in real-world scenarios require stronger generalization abilities of face forgery detectors.
Lip reading in the wild
J. S. Chung and A. Zisserman · 2016
Earlier work this paper cites.
Out of time: Automated lip sync in the wild
Joon Son Chung and Andrew Zisserman · 2016
Earlier work this paper cites.
How far are we from solving the 2d & 3d face alignment problem? (and a dataset of 230,000 3d facial landmarks)
Adrian Bulat and Georgios Tzimiropoulos · 2017
Earlier work this paper cites.
Synthesizing obama: learning lip sync from audio
Supasorn Suwajanakorn, Steven M Seitz, and Ira Kemelmacher-Shlizerman · 2017
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman · 2018
Earlier work this paper cites.
Looking to listen at the cocktail party
Ariel Ephrat, Inbar Mosseri, Oran Lang, Tali Dekel, Kevin W. Wilson, Avinatan Hassidim, William T. Freeman, and Michael Rubinstein · 2018
Earlier work this paper cites.
Deep video portraits
Hyeongwoo Kim, Pablo Garrido, Ayush Tewari, Weipeng Xu, Justus Thies, Matthias Niessner, Patrick Pérez, Christian Richardt, Michael Zollhöfer, and Christian Theobalt · 2018
Earlier work this paper cites.
Cross and learn: Cross-modal self-supervision
Nawid Sayed, Biagio Brattoli, and Björn Ommer · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aäron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Earlier work this paper cites.
Retinaface: Single-stage dense face localisation in the wild
Jiankang Deng, J. Guo, Y. Zhou, Jinke Yu, I. Kotsia, and S. Zafeiriou · 2019
Earlier work this paper cites.
Ms-tcn: Multi-stage temporal convolutional network for action segmentation
Yazan Abu Farha and Juergen Gall · 2019
Earlier work this paper cites.
Exposing deepfake videos by detecting face warping artifacts
Yuezun Li and Siwei Lyu · 2019
Earlier work this paper cites.
Multi-task learning for detecting and segmenting manipulated facial images and videos
H. H. Nguyen, F. Fang, J. Yamagishi, and I. Echizen · 2019
Earlier work this paper cites.
Fsgan: Subject agnostic face swapping and reenactment
Yuval Nirkin, Yosi Keller, and Tal Hassner · 2019
Earlier work this paper cites.
Faceforensics++: Learning to detect manipulated facial images
Andreas Rossler, Davide Cozzolino, Luisa Verdoliva, Christian Riess, Justus Thies, and Matthias Nießner · 2019
Earlier work this paper cites.
Recurrent convolutional strategies for face manipulation detection in videos
Ekraam Sabir, Jiaxin Cheng, Ayush Jaiswal, Wael AbdAlmageed, Iacopo Masi, and Prem Natarajan · 2019
Earlier work this paper cites.
Deferred neural rendering: Image synthesis using neural textures
Justus Thies, Michael Zollhöfer, and Matthias Nießner · 2019
Cited alongside, same era.
Detecting deep-fake videos from phoneme-viseme mismatches
Shruti Agarwal, Hany Farid, Ohad Fried, and Maneesh Agrawala · 2020
Cited alongside, same era.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Henry Zhou, Abdel rahman Mohamed, and Michael Auli · 2020
Cited alongside, same era.
What makes fake images detectable? understanding properties that generalize
Lucy Chai, David Bau, Ser-Nam Lim, and Phillip Isola · 2020
Cited alongside, same era.
Simswap: An efficient framework for high fidelity face swapping
Renwang Chen, Xuanhong Chen, Bingbing Ni, and Yanhao Ge · 2020
Cited alongside, same era.
Learning to recognize patch-wise consistency for deepfake detection
Tianchen Zhao, Xiang Xu, Mingze Xu, Hui Ding, Yuanjun Xiong, and Wei Xia · 2020
Later among the works it cites.
Local relation learning for face forgery detection
Shen Chen, Taiping Yao, Yang Chen, Shouhong Ding, Jilin Li, and Rongrong Ji · 2021
Later among the works it cites.
An empirical study of training self-supervised vision transformers
Xinlei Chen, Saining Xie, and Kaiming He · 2021
Later among the works it cites.
Tclr: Temporal contrastive learning for video representation
Ishan Dave, Rohit Gupta, Mamshad Nayeem Rizve, and Mubarak Shah · 2021
Later among the works it cites.
A large-scale study on unsupervised spatiotemporal representation learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey E. Hinton · 2020
Cited alongside, same era.
The deepfake detection challenge (dfdc) dataset
Brian Dolhansky, Joanna Bitton, Ben Pflaum, Jikuo Lu, Russ Howes, Menglin Wang, and Cristian Canton Ferrer · 2020
Cited alongside, same era.
Deeperforensics-1.0: A large-scale dataset for real-world face forgery detection
Liming Jiang, Wayne Wu, Ren Li, Chen Qian, and Chen Change Loy · 2020
Cited alongside, same era.
Head2head: Video-based neural head synthesis
Mohammad Rami Koujan, Michail Christos Doukas, Anastasios Roussos, and Stefanos Zafeiriou · 2020
Cited alongside, same era.
Face x-ray for more general face forgery detection
Lingzhi Li, Jianmin Bao, Ting Zhang, Hao Yang, Dong Chen, Fang Wen, and Baining Guo · 2020
Cited alongside, same era.
Celeb-df: A large-scale challenging dataset for deepfake forensics
Yuezun Li, Xin Yang, Pu Sun, Honggang Qi, and Siwei Lyu · 2020
Cited alongside, same era.
Two-branch recurrent network for isolating deepfakes in videos
Iacopo Masi, Aditya Killekar, Royston Marian Mascarenhas, Shenoy Pratik Gurudatt, and Wael AbdAlmageed · 2020
Cited alongside, same era.
Christoph Feichtenhofer, Haoqi Fan, Bo Xiong, Ross B. Girshick, and Kaiming He · 2021
Later among the works it cites.
Lips don’t lie: A generalisable and robust approach to face forgery detection
Alexandros Haliassos, Konstantinos Vougioukas, Stavros Petridis, and Maja Pantic · 2021
Later among the works it cites.
Spatial-phase shallow learning: rethinking face forgery detection in frequency domain
Honggu Liu, Xiaodan Li, Wenbo Zhou, Yuefeng Chen, Yuan He, Hui Xue, Weiming Zhang, and Nenghai Yu · 2021
Later among the works it cites.
Towards practical lipreading with distilled and efficient models
Pingchuan Ma, Brais Martinez, Stavros Petridis, and Maja Pantic · 2021
Later among the works it cites.
Contrastive learning of global and local audio-visual representations
Shuang Ma, Zhaoyang Zeng, Daniel McDuff, and Yale Song · 2021
Later among the works it cites.
Audio-visual instance discrimination with cross-modal agreement
Pedro Morgado, Nuno Vasconcelos, and Ishan Misra · 2021
Later among the works it cites.
Videomoco: Contrastive video representation learning with temporally adversarial examples
Tian Pan, Yibing Song, Tianyu Yang, Wenhao Jiang, and Wei Liu · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
M2tr: Multi-modal multi-scale transformers for deepfake detection
Junke Wang, Zuxuan Wu, Jingjing Chen, and Yu-Gang Jiang · 2021
Later among the works it cites.
Multi-attentional deepfake detection
Hanqing Zhao, Wenbo Zhou, Dongdong Chen, Tianyi Wei, Weiming Zhang, and Nenghai Yu · 2021
Later among the works it cites.
One shot face swapping on megapixels
Yuhao Zhu, Qi Li, Jian Wang, Cheng-Zhong Xu, and Zhenan Sun · 2021
Later among the works it cites.
Crossclr: Cross-modal contrastive learning for multi-modal video representations
Mohammadreza Zolfaghari, Yi Zhu, Peter V. Gehler, and Thomas Brox · 2021
Later among the works it cites.