Fetching the paper…
Reading the bibliography…
Multimodal large language models (MLLMs) have demonstrated strong capabilities in vision-related tasks, capitalizing on their visual semantic comprehension and reasoning capabilities.
In Proceedings of the IEEE sixth International Conference on Biometrics: theory, applications and systems
J. Komulainen, A. Hadid & M. Pietikäinen: Context Based Face Anti-Spoofing · 2013
Earlier work this paper cites.
IEEE Transactions on Information Forensics and Security
I. Chingovska, A. R. Anjos & S. Marcel: Biometrics evaluation under spoofing attacks · 2014
Earlier work this paper cites.
In Proceedings of the IEEE International Conference on Image Processing
Z. Boulkenafet, J. Komulainen & A. Hadid: Face Anti-Spoofing Based on Color Texture Analysis · 2015
Earlier work this paper cites.
IEEE Transactions on Information Theory
K. Patel, H. Han & A. K. Jain: Secure face unlock: Spoof detection on smartphones · 2016
Earlier work this paper cites.
IEEE Signal Processing Letters
Z. Boulkenafet, J. Komulainen & A. Hadid: Face antispoofing using speeded-up robust features and fisher vector encoding · 2017
Earlier work this paper cites.
In Proceedings of the IEEE International Joint Conference on Biometrics
Y. Atoum, Y. Liu, A. Jourabloo & X. Liu: Face anti-spoofing using patch and depth-based CNNs · 2017
Earlier work this paper cites.
In Proceedings of the 2nd International Conference on Multimedia and Image Processing
J. Gan, S. Li, Y. Zhai & C. Liu: 3D convolutional neural network based on face anti-spoofing · 2017
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Y. Liu, A. Jourabloo & X. Liu: Learning Deep Models for Face Anti-Spoofing: Binary or Auxiliary Supervision · 2018
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
X. Yang, W. Luo, L. Bao, Y. Gao, D. Gong, S. Zheng, Z. Li & W. Liu: Face anti-spoofing: Model matters, so does data · 2019
Earlier work this paper cites.
In Proceedings of the International Conference on Biometrics
A. George & S. Marcel: Deep pixel-wise binary supervision for face presentation attack detection · 2019
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
A. R"ossler, D. Cozzolino, L. Verdoliva, C. Riess, J. Thies & M. Nießner: Faceforensics++: Learning to detect manipulated facial images · 2019
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
R. Shao, X. Lan, J. Li & P. C. Yuen: Multi-adversarial discriminative deep domain generalization for face presentation attack detection · 2019
Earlier work this paper cites.
In Proceedings of the AAAI Conference on Artificial Intelligence
Y. Qin, C. Zhao, X. Zhu, Z. Wang, Z. Yu, T. Fu, F. Zhou, J. Shi & Z. Lei: Learning meta model for zero-and few-shot face anti-spoofing · 2020
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
G. Wang, H. Han, S. Shan & X. Chen: Cross-domain face presentation attack detection via multi-domain disentangled representation learning · 2020
Earlier work this paper cites.
In Proceedings of the 16th European Conference on Computer Vision
Y. Qian, G. Yin, L. Sheng, Z. Chen & J. Shao: Thinking in frequency: face forgery detection by mining frequency-aware clues · 2020
Earlier work this paper cites.
In Proceedings of the 37th International Conference on Machine Learning
J. Frank, T. Eisenhofer, L. Schönherr, A. Fischer, D. Kolossa & T. Holz: Leveraging frequency analysis for deep fake image recognition · 2020
Earlier work this paper cites.
In A. Vedaldi, H. Bischof, T. Brox & Jan-Michael Frahm.(Eds.): Proceedings of the 16th European Conference on Computer Vision
I. Masi, A. Killekar, R. Marian Mascarenhas, S. P. Gurudatt & W. AbdAlmageed: Two-branch recurrent network for isolating deepfakes in videos · 2020
Earlier work this paper cites.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
L. Li, J. Bao, T. Zhang, H. Yang, D. Chen, F. Wen & B. Guo: Face x-ray for more general face forgery detection · 2020
Earlier work this paper cites.
IEEE Transactions on Information Forensics and Security
A. George, Z. Mostaani, D. Geissenbuhler, O. Nikisins, A’e Anjos & S’ebastien Marcel: Biometric face presentation attack detection with multi-channel convolutional neural network · 2020
Earlier work this paper cites.
Pattern Recognition
Z. Shang, H. Xie, Z. Zha, L. Yu, Y. Li & Y. Zhang: PRRNet: Pixel-region relation network for face forgery detection · 2021
Earlier work this paper cites.
IEEE Transactions on Pattern Analysis and Machine Intelligence
Z. Yu, J. Wan, Y. Qin, X. Li, S. Z. Li & G. Zhao: NAS-FAS: static-dynamic central difference network search for face anti-spoofing · 2021
Earlier work this paper cites.
In Proceedings of the 9th International Conference on Learning Representations
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly & et al: An image is worth 16x16 words: transformers for image recognition at scale · 2021
Earlier work this paper cites.
In Proceedings of the IEEE International Joint Conference on Biometrics
A. George & S’ebastien Marcel: On the effectiveness of vision transformers for zero-shot face anti-spoofing · 2021
Earlier work this paper cites.
IEEE Signal Processing Letters
Z. Yu, X. Li, P. Wang & G. Zhao: TransRPPG: Remote photoplethysmography transformer for 3d mask face presentation attack detection · 2021
Cited alongside, same era.
In Proceedings of the AAAI Conference on Artificial Intelligence
S. Chen, T. Yao, Y. Chen, S. Ding, J. Li & R. Ji: Local relation learning for face forgery detection · 2021
Cited alongside, same era.
In Proceedings of the IEEE/CVF Con- ference on Computer Vision and Pattern Recognition
A. Haliassos, K. Vougioukas, S. Petridis & M. Pantic: Lips Don’t Lie: A generalisable and robust approach to face forgery detection · 2021
Cited alongside, same era.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
H. Zhao, W. Zhou, D. Chen, T. Wei, W. Zhang & N. Yu: Multi-attentional deepfake detection · 2021
Cited alongside, same era.
arXiv preprint arXiv:2111.08687
J. Shao, S. Chen, Y. Li, K. Wang, Z. Yin, Y. He, J. Teng, Q. Sun & et al: Intern: A new learning paradigm towards general vision · 2021
Cited alongside, same era.
In A. Krause, E. Brunskill, K. Cho, B. Engelhardt, S. Sabato & J. Scarlett.(Eds.): International Conference on Machine Learning
J. Li, D. Li, S. Savarese & S. Hoi: BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models · 2023
Later among the works it cites.
In A. Oh, T. Naumann, A. Globerson, K. Saenko, M. Hardt & S. Levine.(Eds.): Proceedings of the 37th International Conference on Neural Information Processing Systems
W. Dai, J. Li, D. Li, A. M. H. Tiong, J. Zhao, W. Wang, B. Li, P. Fung & S. Hoi: InstructBLIP: Towards general-purpose vision-language models with instruction tuning · 2023
Later among the works it cites.
arXiv preprint arXiv:2303.08774
OpenAI: GPT-4 technical report · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.10592
D. Zhu, J. Chen, X. Shen, X. Li & M. Elhoseiny: MiniGPT-4: Enhancing vision-language understanding with advanced large language models · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
IEEE Transactions on Pattern Analysis and Machine Intelligence
Y. Qin, Z. Yu, L. Yan, Z. Wang, C. Zhao & Z. Lei: Meta-teacher for face anti-spoofing · 2022
Cited alongside, same era.
In Proceedings of the IEEE International Conference on Image Processing
Z. Ming, Z. Yu, M. Al-ghadi, M. Visani, M. M. Luqman & Jean-christophe Burie: Vitranspad: video transformer using convolution and self-attention for face presentation attack detection · 2022
Cited alongside, same era.
IEEE Transactions on Biometrics, Behavior, and Identity Science
Z. Wang, Q. Wang, W. Deng & G. Guo: Face anti-spoofing using transformers with relation-aware mechanism · 2022
Cited alongside, same era.
arXiv preprint arXiv:2304.07549
A. Liu & Y. Liang: MA-vit: modality-agnostic vision transformers for face anti-spoofing · 2022
Cited alongside, same era.
IEEE Transactions on Information Theory
Z. Wang, Q. Wang, W. Deng & G. Guo: Learning multi-granularity temporal characteristics for face anti-spoofing · 2022
Cited alongside, same era.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
J. Cao, C. Ma, T. Yao, S. Chen, S. Ding & X. Yang: End-to-end reconstruction-classification learning for face forgery detection · 2022
Cited alongside, same era.
In Proceedings of the AAAI Conference on Artificial Intelligence
Q. Gu, S. Chen, T. Yao, Y. Chen, S. Ding & R. Yi: Exploiting fine-grained face forgery clues via progressive enhancement learning · 2022
Cited alongside, same era.
J. Bai, S. Bai, S. Yang, S. Wang, S. Tan, P. Wang, J. Lin, C. Zhou & J. Zhou: Qwen-vl: A frontier large vision-language model with versatile abilities · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.08485
H. Liu, C. Li, Q. Wu & Y. J. Lee: Visual instruction tuning · 2023
Later among the works it cites.
arXiv preprint arXiv:2306.13394
C. Fu, P. Chen, Y. Shen, Y. Qin, M. Zhang, X. Lin, Z. Qiu, W. Lin, J. Yang, X. Zheng, K. Li, X. Sun & R. Ji: MME: A comprehensive evaluation benchmark for multimodal large language models · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.14178
Q. Ye, H. Xu, G. Xu, J. Ye, M. Yan, Y. Zhou, J. Wang, A. Hu & et al: MPLUG-owl: Modularization empowers large language models with multimodality · 2023
Later among the works it cites.
arXiv preprint arXiv:2311.17092
B. Li, Y. Ge, Y. Ge, G. Wang, R. Wang, R. Zhang & Y. Shan: Seed-ench-2: benchmarking multimodal large language models · 2023
Later among the works it cites.
arXiv preprint arXiv:2312.02896
R. Cai, Z. Song, D. Guan, Z. Chen, X. Luo, C. Yi & A. C. Kot: Benchlmm: Benchmarking cross-style visual capability of large multimodal models · 2023
Later among the works it cites.
arXiv preprint arXiv:2309.14181
H. Wu, Z. Zhang, E. Zhang, C. Chen, L. Liao, A. Wang, C. Li, W. Sun & et al: Q-bench: A benchmark for general-purpose foundation models on low-level vision · 2023
Later among the works it cites.
arXiv preprint arXiv:2309.02218
H. Song, S. Huang, Y. Dong & W.-W. Tu: Robustness and generalizability of deepfake detection: A study with diffusion models · 2023
Later among the works it cites.
arXiv preprint arXiv:2311.09193
Y. Wu, P. Zhang, W. Xiong, B. Oguz, J. C. Gee & Y. Nie: The role of chain-of-thought in complex vision-language reasoning task · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.08485
H. Liu, C. Li, Q. Wu & Y J. Lee:Visual instruction tuning · 2023
Later among the works it cites.
arXiv preprint arXiv:2303.08774
A. Josh, A. Steven, A. Sandhini, A. Lama, A. Ilge, A. Florencia Leoni, A. Diogo, A. Janko, A. Sam, A. Shyamal & et al: GPT-4 technical report · 2023
Later among the works it cites.
In Proceedings of the IEEE/CVF International Conference on Computer Vision
K. Srivatsan, M. Naseer & K. Nandakumar: Flip: cross-domain face anti-spoofing with language guidance · 2023
Later among the works it cites.
arXiv preprint arXiv:2303.12712
Q. Dong, L. Li, D. Dai, C. Zheng, Z. Wu, B. Chang, X. Sun, J. Xu & Z. Sui: A survey on in-context learning · 2023
Later among the works it cites.
arXiv preprint arXiv:2301.00234
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. Lee, Y. Li, S. Lundberg & et al: Sparks of artificial general intelligence: Early experiments with GPT-4 · 2023
Later among the works it cites.
IEEE Transactions on Dependable and Secure Computing
Z. Yu, R. Cai, Z. Li, W. Yang, J. Shi & A. C. Kot: Benchmarking joint face spoofing and forgery detection with visual and physiological cues · 2024
Closest in time.
IEEE Transactions on Multimedia
M. Fang, L. Yu, Y. Song, Y. Zhang & H. Xie: IEIRNet: Inconsistency exploiting based identity rectification for face forgery detection · 2024
Closest in time.
IEEE/ACM Transactions on Audio, Speech, and Language Processing
L. Wang, L. Yu, Y. Zhang & H. Xie: Generalizable speech spoofing detection against silence trimming with data augmentation and multi-task meta-learning · 2024
Closest in time.
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
T. Guan, F. Liu, X. Wu, R. Xian, Z. Li, X. Liu, X. Wang, L. Chen,F. Huang,Y. Yacoob & et al: Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models · 2024
Closest in time.
arXiv preprint arXiv:2401.08276
Y. Huang, Q. Yuan, X. Sheng, Z. Yang, H. Wu, P. Chen, Y. Yang, L. Li & W. Lin: Aesbench: An expert benchmark for multimodal large language models on image aesthetics perception · 2024
Closest in time.