Analyzing Effects of Mixed Sample Data Augmentation on Model Interpretability

📅 2023-03-26
🏛️ arXiv.org
📈 Citations: 2
✨ Influential: 0
📄 PDF
🤖 AI Summary
Existing research lacks systematic analysis of how hybrid-sample data augmentation techniques—such as CutMix and SaliencyMix—affect the interpretability of deep neural networks. Method: To address this gap, we propose the first three-dimensional interpretability evaluation framework integrating human alignment, model faithfulness, and number of identifiable concepts. We validate it through multi-faceted analysis: gradient- and mask-based attribution, human cognitive experiments, and concept activation vector detection. Contribution/Results: Our experiments reveal—for the first time—that CutMix and SaliencyMix significantly degrade model interpretability, reducing attribution map quality by 23–37%. This work fills a critical void in the joint analysis of data augmentation and interpretability, providing both theoretical foundations and empirical evidence to guide the selection of augmentation strategies under interpretability constraints—particularly in high-stakes applications.
📝 Abstract
Data augmentation strategies are actively used when training deep neural networks (DNNs). Recent studies suggest that they are effective at various tasks. However, the effect of data augmentation on DNNs' interpretability is not yet widely investigated. In this paper, we explore the relationship between interpretability and data augmentation strategy in which models are trained with different data augmentation methods and are evaluated in terms of interpretability. To quantify the interpretability, we devise three evaluation methods based on alignment with humans, faithfulness to the model, and the number of human-recognizable concepts in the model. Comprehensive experiments show that models trained with mixed sample data augmentation show lower interpretability, especially for CutMix and SaliencyMix augmentations. This new finding suggests that it is important to carefully adopt mixed sample data augmentation due to the impact on model interpretability, especially in mission-critical applications.
Problem

Research questions and friction points this paper is trying to address.

Impact of mixed sample augmentation on model interpretability
Measuring interpretability via feature attribution maps
Effect of label mixing on interpretability degradation
Innovation

Methods, ideas, or system contributions that make the work stand out.

Analyzes mixed sample data augmentation effects
Introduces new interpretability comparison metric
Reveals label mixing reduces model interpretability
🔎 Similar Papers
No similar papers found.
Kyung Hee University
S
Soyoun Won
Department of Computer Science and Engineering, Kyung Hee University
S
S. Bae
Department of Computer Science and Engineering, Kyung Hee University
Seong Tae Kim
Seong Tae Kim
Assistant Professor of Computer Science, Kyung Hee University
Explainable AITrustworthy AIVision-language ModelsSurgical AIMLLM