cs.AIOct 7, 2026

Open-MMUnlearning: Unifying Methods and Evaluation for MLLM Unlearning

Authors: Junkai Chen, Yuhao He, Qianshan Wei, Junxiang You, Jingwen Shao, Junkai Lin, Zhongkai Yue, Xiaotian Ye, +14 more

Organizations: Institute of Automation, Chinese Academy of Sciences

Abstract

As multimodal large language models (MLLMs) become more capable and widely deployed, concerns about privacy and safety have become increasingly pressing. Machine unlearning offers one approach to addressing these concerns by removing designated information from trained models while preserving unrelated capabilities. However, fragmented implementations and evaluation protocols, incomplete robustness testing, and limited understanding of metric reliability make progress in MLLM unlearning difficult to assess systematically. We introduce Open-MMUnlearning, an open-source, extensible framework that integrates target-model preparation, multimodal data processing, unlearning, and evaluation through shared interfaces and structured configurations. The framework supports five benchmarks spanning privacy, safety, and copyright, eight MLLMs from four model families, and twelve unlearning methods. Its evaluation suite jointly assesses forgetting effectiveness, retained utility, and robustness to model interventions, adversarial inputs, and membership inference attacks. Using a common evaluation protocol, we compare ten representative unlearning methods. In this comparison, GD and MIP-Editor tie for the highest overall score: GD achieves the highest Forget Quality, while MIP-Editor preserves more Model Utility. We further introduce a metric meta-evaluation protocol that tests faithfulness using models with controlled exposure to target knowledge and robustness under quantization and relearning. Among the thirteen evaluated metrics, BLEU achieves the highest aggregate reliability score. KS-Test attains the highest faithfulness AUC but performs less well on robustness. Together, the framework and these findings support reproducible comparison of MLLM unlearning methods and systematic assessment of evaluation reliability.

Figures & tables

Appendix figures & tables15 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Exploring and Bridging Knowledge Holes in Unlearned Multimodal Large Language Models

    Aug 3, 2026Junxiang You, Junkai Chen, Yuhao He +3Multimodal Large Language ModelsMLLM Unlearning

  2. MLUBench: A Benchmark for Lifelong Unlearning Evaluation in MLLMs

    Jun 11, 2026He Li, Haoang Chi, Qizhou Wang +6VLM UnlearningMultimodal Large Language Models

  3. Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models

    Aug 2, 2026Junkai Lin, Junkai Chen, Siqi Hou +5VLM UnlearningPrivacy Leakage in Language Models