OODFace: Benchmarking Robustness of Face Recognition under Common Corruptions and Appearance Variations

📅 2024-12-03
🏛️ arXiv.org
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
Existing face recognition models exhibit insufficient robustness under Out-of-Distribution (OOD) conditions, particularly degrading significantly when confronted with real-world image degradations and appearance variations. To address this, we introduce OODFace—the first OOD-robustness benchmark specifically designed for face verification—comprising 30 common degradation types (e.g., noise, blur, compression) and appearance variations (e.g., pose, illumination, masks, makeup), and establishing three evaluation splits: LFW-C/V, CFP-FP-C/V, and YTF-C/V. We propose the first dual-dimensional (degradation + appearance) modeling framework for facial OOD challenges, along with a scalable, unified evaluation toolkit. Extensive experiments reveal that 19 open-source models and 3 commercial APIs suffer over 40% accuracy drops under occlusion, illumination shifts, and mask perturbations. We further validate the efficacy of vision-language model–assisted analysis and physical-world experiments. All code, data, and evaluation tools are publicly released.

Technology Category

Computer Vision: Adversarial Attacks & RobustnessMachine Learning: Adversarial Learning & RobustnessNatural Language Processing: Fact-Checking / Misinformation Detection (NLP Focus)

Application Category

User Modeling, Personalization and Recommendation: Fairness-aware retrieval and rankingSearch and Retrieval-Augmented AI: Web evaluation methodologies and metricsEconomics, Online Markets and Human Computation: Data quality aspects of human-annotated datasets
📝 Abstract
With the rise of deep learning, facial recognition technology has seen extensive research and rapid development. Although facial recognition is considered a mature technology, we find that existing open-source models and commercial algorithms lack robustness in certain complex Out-of-Distribution (OOD) scenarios, raising concerns about the reliability of these systems. In this paper, we introduce OODFace, which explores the OOD challenges faced by facial recognition models from two perspectives: common corruptions and appearance variations. We systematically design 30 OOD scenarios across 9 major categories tailored for facial recognition. By simulating these challenges on public datasets, we establish three robustness benchmarks: LFW-C/V, CFP-FP-C/V, and YTF-C/V. We then conduct extensive experiments on 19 facial recognition models and 3 commercial APIs, along with extended physical experiments on face masks to assess their robustness. Next, we explore potential solutions from two perspectives: defense strategies and Vision-Language Models (VLMs). Based on the results, we draw several key insights, highlighting the vulnerability of facial recognition systems to OOD data and suggesting possible solutions. Additionally, we offer a unified toolkit that includes all corruption and variation types, easily extendable to other datasets. We hope that our benchmarks and findings can provide guidance for future improvements in facial recognition model robustness.
Problem

Research questions and friction points this paper is trying to address.

Assessing face recognition robustness in OOD scenarios
Exploring common corruptions and appearance variations challenges
Providing benchmarks and solutions for model reliability
Innovation

Methods, ideas, or system contributions that make the work stand out.

Introduces 30 OOD scenarios for robustness testing
Uses defense strategies and Vision-Language Models
Provides unified toolkit for corruption and variation
🔎 Similar Papers
2024-01-21IEEE International Conference on Automatic Face & Gesture RecognitionCitations: 4
💼 Related Jobs
No related jobs found.
Beihang University | China Academy of Information and Communications Technology
C
Cai Kang
Institute of Artificial Intelligence, Beihang University
Yubo Chen
Yubo Chen
Institute of Automation, Chinese Academy of Sciences
Natural Language ProcessingInformation ExtractionEvent ExtractionLarge Language Model
S
Shouwei Ruan
Institute of Artificial Intelligence, Beihang University
Shiji Zhao
Shiji Zhao
Beihang University
Machine LearningTrustworthy AIExplainable AIRobust AI
Ruochen Zhang
Ruochen Zhang
Brown University
Multilingual NLPInterpretabilityCode-Switching
J
Jiayi Wang
China Academy of Information and Communications Technology
S
Shan Fu
China Academy of Information and Communications Technology
Xingxing Wei
Xingxing Wei
Professor of Artificial Intelligence, Beihang University
Computer visionAdversarial machine learning