Purifying LoRA-Tuned LLMs via Null-Space Projection
October 2, 2026
Null-space projection removes backdoor triggers from LoRA-tuned models without requiring access to clean references, prior trigger knowledge, or post-hoc retraining. The method reduces attack success rates while preserving both the base model's general capabilities and the downstream skills learned by the adapter.
HOW THIS AFFECTS YOU
●
builderYou can secure PEFT-based deployments against poisoning attacks without expensive retraining cycles.
●
policyThis provides a technical path for mitigating supply chain risks in fine-tuned models.