Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

FM-ViT: Flexible Modal Vision Transformers for Face Anti-Spoofing (2305.03277v1)

Published 5 May 2023 in cs.CV

Abstract: The availability of handy multi-modal (i.e., RGB-D) sensors has brought about a surge of face anti-spoofing research. However, the current multi-modal face presentation attack detection (PAD) has two defects: (1) The framework based on multi-modal fusion requires providing modalities consistent with the training input, which seriously limits the deployment scenario. (2) The performance of ConvNet-based model on high fidelity datasets is increasingly limited. In this work, we present a pure transformer-based framework, dubbed the Flexible Modal Vision Transformer (FM-ViT), for face anti-spoofing to flexibly target any single-modal (i.e., RGB) attack scenarios with the help of available multi-modal data. Specifically, FM-ViT retains a specific branch for each modality to capture different modal information and introduces the Cross-Modal Transformer Block (CMTB), which consists of two cascaded attentions named Multi-headed Mutual-Attention (MMA) and Fusion-Attention (MFA) to guide each modal branch to mine potential features from informative patch tokens, and to learn modality-agnostic liveness features by enriching the modal information of own CLS token, respectively. Experiments demonstrate that the single model trained based on FM-ViT can not only flexibly evaluate different modal samples, but also outperforms existing single-modal frameworks by a large margin, and approaches the multi-modal frameworks introduced with smaller FLOPs and model parameters.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (10)
  1. Ajian Liu (31 papers)
  2. Zichang Tan (25 papers)
  3. Zitong Yu (119 papers)
  4. Chenxu Zhao (29 papers)
  5. Jun Wan (79 papers)
  6. Yanyan Liang (29 papers)
  7. Zhen Lei (205 papers)
  8. Du Zhang (9 papers)
  9. Stan Z. Li (222 papers)
  10. Guodong Guo (75 papers)
Citations (34)

Summary

We haven't generated a summary for this paper yet.