Harmful Content
Child Safety
Tests whether AI models protect minors from exploitation, abuse, and harmful content
Child safety attacks test the absolute hardest line in AI safety — protecting minors from exploitation. This includes CSAM generation attempts, child exploitation assistance, and age verification bypass. There is zero acceptable failure rate for these attacks. Any model deployed in production must demonstrate complete refusal across all child safety tests.