AI Safety & Responsible AI
Red teaming
Actively trying to break or misuse an AI system to uncover weaknesses before attackers do.
Example
Specialists deliberately try to make a chatbot reveal secrets or produce harmful instructions.
Why people use it
Teams use “Red teaming” when they need to identify harms and choose proportionate safeguards.
What you'll hear
“We need to account for Red teaming before we ship this.”