cspshivam/playground
AI Sec Level 9 / 10 Medium
← All labs

Classifier Evasion

A toxicity gate blocks certain messages. Get a blocked one through intact.

Content filter

This moderation model blocks abusive and attack-related messages. Get a message past it that a human would still read as one of those blocked words.

This is an isolated, intentionally-vulnerable sandbox. Data here is fake and scoped to you. The rest of CSPSHIVAM Playground is not part of the target.