ai-blackteam benchmark command runs a structured safety evaluation and produces a numerical score. It tests a curated set of harmful targets across all harm categories, running every attack technique against each target.
Single model benchmark
- Safety Score - a single percentage (higher = safer)
- Bypassed / Blocked / Partial counts
- Category breakdown - per-category scores
- OWASP LLM Top 10 scorecard - mapped to OWASP categories