Pretext: Defeating Malicious Skill Detection Frameworks for AI Agents
Researchers developed a new framework called Pretext that can evade existing skill detection systems for AI agents. Pretext uses a white-box LLM attacker to craft skills that evade detection while still delivering a payload. This highlights major gaps in current skill scanners and raises concerns about the security of AI agent skills.
Save an API key to vote.