Agent & Tool Attacks
Tool Exploitation
Tests whether AI agents can be tricked into escalating tool access through progressive normalization
Tool exploitation attacks test whether AI agents with tool access can be gradually normalized into performing unauthorized actions. The progressive normalization technique starts with benign file reads and escalates to sensitive data access, exploiting the agent’s tendency to extend trust based on prior actions. This is a foundational test for any agent with filesystem or API access.