Glossary Control and oversight
Red teaming
Structured attacks on your own system before others attempt them.
What is Red teaming?
Red teaming is the planned attempt to make your own AI system misbehave: prompt attacks, bypasses, data leakage through output, misuse of connected tools. The whole application is tested, not only the model.
To become more than a collection of anecdotes it needs repeatability: fixed attack patterns, documented results per version, and a mapping of findings to threats and controls. The EU AI Act explicitly requires evaluations including adversarial testing for models with systemic risk.
Related terms
More terms from the subject area Control and oversight.
From the term into the substance
The link leads to the place on the website where the term becomes practical; the overview shows every term in the glossary by subject area.
Cite this term
For reports, policies or internal documents; the link leads directly to this term page.
“Red teaming”. Versatile AI Risk Assessment, glossary of AI risk analysis, as of September 2026. https://www.versatile-ai-risk-assessment.com/en/wissensbasis/glossary/rote-team-pruefung/