Divide and Inject: Can Agents Reconstruct an Indirect Prompt Injection from Fragments?

This paper introduces a new attack method called adaptive long-context prompt injection (AdaLCPI) that reconstructs malicious objectives from incomplete fragments, which can be used to compromise AI agents. This highlights the need for robust safety evaluations of agents against such attacks.

RSS Score 0 9/30/2026, 4:00:00 AM Original Source
Save an API key to vote.