Faithful Dual-constrained Erasure for Robust LLM Safety Alignment
Researchers propose FDCU, a novel framework to enforce authentic memory deletion in Large Language Models to prevent retraining attacks. FDCU restricts parameter updates and promotes the dismantling of target representations.
Save an API key to vote.