Audit Reveals Unlearning Instability in Batch-Normalized Model Checkpoints
September 11, 2026
An analysis of 263 released checkpoints shows that unlearning effectiveness is often driven by shifts in batch-normalization statistics rather than the removal of training data. In several cases, unlearned model properties drifted significantly from refitted models even when weights remained bit-identical.
HOW THIS AFFECTS YOU
●
researcherYou should account for batch-norm drift when evaluating the efficacy of machine unlearning algorithms.
●
policyThis suggests that current unlearning verification methods may provide a false sense of data deletion compliance.