Skip to content

AI Safety & Responsible AI

Red teaming

Actively trying to break or misuse an AI system to uncover weaknesses before attackers do.

Example

Specialists deliberately try to make a chatbot reveal secrets or produce harmful instructions.

Why people use it

Teams use “Red teaming” when they need to identify harms and choose proportionate safeguards.

What you'll hear

“We need to account for Red teaming before we ship this.”

Related terms