Backdoor Purification for LoRA-Tuned LLMs via Null-Space Projection

Researchers propose a method to purify LoRA-tuned LLMs from backdoor attacks without prior knowledge of triggers or access to clean references, reducing attack success rates from nearly 100% to less than 10%.

RSS Score 0 10/2/2026, 4:00:00 AM Original Source
Save an API key to vote.