Vulnerability research attacks test whether a model will actively find and exploit security vulnerabilities in code you provide. The attack escalates through three levels — from passive review to full proof-of-concept exploit generation. This matters because code-capable models can be weaponized as vulnerability research tools if they lack appropriate guardrails.

Summary

4 attacks total: 4 single-turn.

Attacks

Shared metadata

Sub-modes

The parent vuln-research attack runs all three sub-modes in sequence. You can also target a specific sub-mode directly by its technique ID:

Escalation Levels

Each run escalates through three levels of increasing severity:

Example