Open-MMUnlearning: Unifying Methods and Evaluation for MLLM Unlearning

📅 2026-10-07
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the absence of standardized evaluation criteria, fragmented implementations, and insufficient robustness verification in machine unlearning for multimodal large language models. To this end, we present an open-source framework integrating data, model, and evaluation modules. Methodologically, we propose a metric meta-evaluation protocol to quantify assessment reliability, support multi-benchmark reproducibility and systematic comparison via unified interfaces, and introduce adversarial attack simulations alongside membership inference detection for robustness quantification. Experimental results demonstrate that Gradient Difference (GD) and MIP-Editor achieve superior performance, while BLEU is validated as a highly reliable evaluation metric. By bridging the gap in standardized comparative pipelines, this work significantly enhances the reproducibility and methodological rigor of unlearning evaluation.
📝 Abstract
As multimodal large language models (MLLMs) become more capable and widely deployed, concerns about privacy and safety have become increasingly pressing. Machine unlearning offers one approach to addressing these concerns by removing designated information from trained models while preserving unrelated capabilities. However, fragmented implementations and evaluation protocols, incomplete robustness testing, and limited understanding of metric reliability make progress in MLLM unlearning difficult to assess systematically. We introduce Open-MMUnlearning, an open-source, extensible framework that integrates target-model preparation, multimodal data processing, unlearning, and evaluation through shared interfaces and structured configurations. The framework supports five benchmarks spanning privacy, safety, and copyright, eight MLLMs from four model families, and twelve unlearning methods. Its evaluation suite jointly assesses forgetting effectiveness, retained utility, and robustness to model interventions, adversarial inputs, and membership inference attacks. Using a common evaluation protocol, we compare ten representative unlearning methods. In this comparison, GD and MIP-Editor tie for the highest overall score: GD achieves the highest Forget Quality, while MIP-Editor preserves more Model Utility. We further introduce a metric meta-evaluation protocol that tests faithfulness using models with controlled exposure to target knowledge and robustness under quantization and relearning. Among the thirteen evaluated metrics, BLEU achieves the highest aggregate reliability score. KS-Test attains the highest faithfulness AUC but performs less well on robustness. Together, the framework and these findings support reproducible comparison of MLLM unlearning methods and systematic assessment of evaluation reliability.
Problem

Research questions and friction points this paper is trying to address.

Multimodal Large Language Models
Machine Unlearning
Evaluation Protocols
Privacy and Safety
Metric Reliability
Innovation

Methods, ideas, or system contributions that make the work stand out.

Multimodal Machine Unlearning
Evaluation Framework
Metric Meta-Evaluation
Robustness Testing
MLLM
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
J
Junkai Chen
Institute of Automation, Chinese Academy of Sciences
Y
Yuhao He
Institute of Automation, Chinese Academy of Sciences
Q
Qianshan Wei
Institute of Automation, Chinese Academy of Sciences
J
Junxiang You
Institute of Automation, Chinese Academy of Sciences, University of the Chinese Academy of Sciences
J
Jingwen Shao
ByteDance
J
Junkai Lin
The Chinese University of Hong Kong
Z
Zhongkai Yue
Institute of Automation, Chinese Academy of Sciences
Xiaotian Ye
Xiaotian Ye
Beijing University of Posts and Telecommunications
Natural Language ProcessingKnowledge RepresentationLarge Language Models
Z
Zhengbo Jiao
Shanghai University of Finance and Economics
Jiali Cheng
Jiali Cheng
UMass Lowell
Trustworthy AILanguage AgentsAI4Science
Zhijie Deng
Zhijie Deng
The Hong Kong University of Science and Technology (Guangzhou)
LLM unlearningTrustworthy LLMMLLM
K
Kening Zheng
University of Illinois at Chicago
Ruiqi Liu
Ruiqi Liu
Texas Tech University
nonparametric methodsmachine learningeconometrics
Hadi Amiri
Hadi Amiri
Assistant Professor, University of Massachusetts Lowell
Natural Language ProcessingHealthcare
Y
Yi Yu
Jilin University
Zhenan Sun
Zhenan Sun
Institute of Automation, Chinese Academy of Sciences
BiometricsPattern RecognitionComputer Vision
Qi Li
Qi Li
Institute of Automation, Chinese Academy of Sciences
pattern recognitioncomputer vision
Ka-Ho Chow
Ka-Ho Chow
The University of Hong Kong
Trustworthy AICybersecurityML for SystemsSystems for ML
Sijia Liu
Sijia Liu
Red Cedar Distinguished Associate Professor, Michigan State Univ.; Affiliate Prof., IBM Research
Adversarial RobustnessMachine UnlearningAI SafetyBlack-box OptimizationSignal Processing
Liang Wang
Liang Wang
Institute of Psychology, Chinese Academy of Sciences
ECoGfMRINeuronal oscillationsBrain networksSpatial attention
J
Jiaqi Li
The University of Hong Kong
S
Shu Wu
Institute of Automation, Chinese Academy of Sciences