← Back to all articles
arXiv cs.LGOctober 2, 2026

Backdoor Purification for LoRA-Tuned LLMs via Null-Space Projection

Excerpt

arXiv:2610.00685v1 Announce Type: cross Abstract: With the rapid adoption of large language models (LLMs) and parameter-efficient fine-tuning (PEFT) methods, the risk of backdoor attacks has become more severe. Existing backdoor purification methods typically rely on at least one of the strong assumptions, such as prior knowledge of triggers, access to clean references, or aggressive retraining, and they often lack comprehensive evaluations. These constraints substantially limit their practical