AI GlossaryㅋSafety and controversy
constitution
An internal rulebook that defines the line between requests an AI should answer and requests it should refuse
In plain words
A constitution is a written set of rules that tells an AI system which questions it's allowed to answer and which ones it needs to block. Think of it like the training manual a company hands to new employees. If the manual only has vague guidelines, employees get confused and end up rejecting even harmless requests just to be safe. AI safety systems work the same way: when the constitution is written poorly, the system starts mistaking normal, safe questions for dangerous ones and blocks them too.
That's why developers go back and rewrite this document, build new training data based on the revised standard, and retrain the program that makes these judgment calls. In effect, they're refining the manual itself so it can tell ambiguous cases apart more accurately.
How it shows up in the news
In the article, it's described as: "they rewrote the constitution — the rule set behind the classifier — rebuilt the training data around it, and retrained the classifier." Here, constitution doesn't mean a national charter; it refers to the internal rulebook that defines what an AI should treat as dangerous.
See also
Stories using this term
- Claude's Fable 5 Cuts Biology False Positives by 85%AI · 2026.08.09
- Claude Fable 5 eases biology block rate by 85%AI · 2026.08.08
- Claude tests 'Morning Brief' feature built on scheduled tasksAI · 2026.08.26
- Claude Plugins Now Installable in Two Commands via GitHub MirrorAI · 2026.08.24
- Claude Code lets teams package a full setup into one pluginAI · 2026.09.04
- Claude MCP Connectors Get Centralized Enterprise AuthenticationAI · 2026.08.25
