๐ค AI Summary
This work addresses the โright to be forgottenโ by studying efficient machine unlearning under privacy compliance: removing the influence of specified data from a trained model without full retraining, such that the post-unlearning model distribution closely approximates that of a model trained de novo on the remaining data. Methodologically, it establishes, for the first time under convexity assumptions, theoretically certified guarantees for approximate unlearning; introduces a dynamic tracking technique based on the infinite Wasserstein distance, yielding provably improved computational complexity; and integrates projected-noise SGD, differential-privacy-inspired noise injection, and convex optimization analysis. Experiments demonstrate that, under both mini-batch and full-batch settings, the proposed method reduces gradient computation cost to just 2% and 10% of the state-of-the-art, respectively, while preserving equivalent model utility and formal privacy guarantees.
๐ Abstract
``The right to be forgotten'' ensured by laws for user data privacy becomes increasingly important. Machine unlearning aims to efficiently remove the effect of certain data points on the trained model parameters so that it can be approximately the same as if one retrains the model from scratch. We propose to leverage projected noisy stochastic gradient descent for unlearning and establish its first approximate unlearning guarantee under the convexity assumption. Our approach exhibits several benefits, including provable complexity saving compared to retraining, and supporting sequential and batch unlearning. Both of these benefits are closely related to our new results on the infinite Wasserstein distance tracking of the adjacent (un)learning processes. Extensive experiments show that our approach achieves a similar utility under the same privacy constraint while using $2%$ and $10%$ of the gradient computations compared with the state-of-the-art gradient-based approximate unlearning methods for mini-batch and full-batch settings, respectively.