Approval Laundering: Systematizing Approval--Execution Binding Failures in AI Coding-Agent Harnesses
Researchers introduce Approval Laundering, a taxonomy of six failure modes in AI coding-agent harnesses that silently substitute one action for another after approval. They evaluate these modes using a controlled study and prototype Approval Token, a capability that eliminates two of the failure modes.
Save an API key to vote.