Skip to content

AI Safety & Responsible AI

Adversarial testing

Testing a system with deliberately difficult or manipulative inputs.

Example

Evaluators test prompts designed to exploit known weaknesses rather than ordinary use.

Why people use it

Knowing how “Adversarial testing” works helps teams identify harms and choose proportionate safeguards.

What you'll hear

“We need to account for Adversarial testing before we ship this.”

Related terms