← Back to the AI Glossary 📖 Safety & Alignment

Red Teaming

Deliberately testing an AI system by trying to make it fail, misbehave, or produce harmful output, in order to find and fix weaknesses before real users do.

AI labs run extensive red-teaming before releasing a new model, and often invite external security researchers to try to break it too.
One of 60 free AI glossary terms
Plain-language definitions for the AI jargon you'll actually run into — no email needed, ever, for this section.
Browse the Full Glossary →