Multi-agent exploitation attacks target systems where multiple AI agents communicate and delegate tasks. Attackers exploit trust relationships between agents through impersonation, session smuggling, delegation abuse, and cascading jailbreaks. When one agent in a chain is compromised, the entire pipeline can be subverted.

Summary

5 attacks total: 4 multi-turn, 1 tool-use.

Attacks

Example