Related Experiment Videos
Dynamic Cross-Modal Modeling with an Ultra-Lightweight Architecture for Face Anti-Spoofing
Nana Li1, Jiayu Wang1, Zuhe Li1
1School of Computer Science and Artificial Intelligence, Zhengzhou University of Light Industry, Zhengzhou 450002, China.
None:
The effectiveness of multimodal face anti-spoofing largely depends on the modeling of cross-modal relationships. However, most existing approaches rely on static fusion or implicitly learned feature aggregation, which assumes fixed modality importance, limiting its ability to capture reliability variations across different attack patterns. Under strict computational constraints, achieving effective dynamic cross-modal modeling remains a significant challenge. To address this issue, we propose an ultra-lightweight dynamic cross-modal framework for face anti-spoofing, with ultra-low parameters, FLOPs, latency, memory and high FPS for real-time edge inference. A compact feature extractor is constructed by enhancing ShuffleNetV2 with the Ghost-Generated Shuffle BlockA (GGS-BlockA), which significantly reduces redundant computation while maintaining high discriminative capability. On this basis, a Lightweight Cross-Modal Attention (LCMA) module performs sample-wise dynamic modality reweighting to capture reliability variations among RGB, Depth, and IR modalities. Furthermore, a Lightweight Cross-Modal Fusion (LCMF) module utilizes depth cues as stable guidance to improve cross-modal feature alignment and complementary representation. Experiments on the CASIA-SURF benchmark demonstrate that the proposed method achieves an Average Classification Error Rate (ACER) of 0.064% with only 0.14M parameters and 0.0065G FLOPs. At the strict threshold of TPR@FPR=10-4, a detection rate of 99.86% is obtained, demonstrating strong robustness and generalization capability under extremely low computational cost.
Related Concept Videos
Masking and Demasking Agents
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on the metal...
Facial Feedback Hypothesis
Modeling and Similitude
Nonconscious Mimicry
Muscles for Facial Expressions
Light Acquisition