AI GlossaryㄹSafety and controversy
Red Teaming
A role that deliberately attacks a product before launch — "we break it ourselves first" has become the standard procedure for AI safety.
In plain words
It's a role, or activity, where people deliberately try to attack a product before it's released. The term comes from military exercises, where a simulated enemy force (the red team) attacks friendly forces (the blue team). It moved through the security industry and has since become standard practice for AI.
Before a new model launches, hundreds of people stress-test it, trying to jailbreak it, provoke dangerous answers, and surface biases, and the results get published in the model card. Phrases like "reviewed by external red teamers for months" have become a trust signal in launch announcements. Government regulations are also increasingly building red-teaming requirements into rules for frontier models.
See also
Stories using this term
- Anthropic finds collaboration breaks down in agent swarm experimentsAI · 2026.08.17
- ChatGPT sites now handled by Codex instead of Git and CI, adds team invitesAI · 2026.08.21
- NVIDIA Details How to Build Fully Isolated K8s Tenants on Shared GPUsAI · 2026.08.09
- PDI Unveils Agentic Internal App Deployment System Built on AWSAI · 2026.08.09
- Kimi launches 100-agent parallel swarm systemAI · 2026.08.08
- The 'Claude Tag' from Slack is coming to Claude DesktopAI · 2026.08.17
