When Tools Silently Lie: Evaluating and Mitigating Blind Compliance in Tool-Augmented Data Agents
Researchers introduce ToxicBench to evaluate and mitigate 'blind compliance' in tool-augmented data agents, which can lead to incorrect evidence and wrong-answer adoption. They measure checking and adoption under various errors and find that poisoning lowers task success by 26-39 percentage points. The study highlights the importance of evidence availability and answer selection in agent reliability.
Save an API key to vote.