AdaPop Method Reduces LLM Unlearning Leakage via Popularity-Weighted Gradients
August 17, 2026
AdaPop scales gradient pressure based on fact popularity using external proxies like Wikidata. The method uses a dual-ascent controller to balance forgetting and retention, reducing content leakage by 5x under paraphrased queries compared to uniform unlearning approaches.
HOW THIS AFFECTS YOU
●
researcherYou can use popularity-dependent exponents to better target deeply memorized facts during unlearning.
●
policyThis provides a more robust method for removing sensitive or copyrighted information from model weights.