METAL for iPhone

Read AI news in the METAL app.

Download METAL and discover fresh AI stories every day.

Download on the App Store

For iPhone · Free download

Search for METAL AI Magazine in the App Store on your iPhone.

METAL

AI GlossaryㅈSafety and controversy

Automated Safety Review

A screening step in which an automated system—not a human—checks API requests for signs of misuse without anyone reading the actual conversation content.

In plain words

Automated Safety Review is a checkpoint where a machine, not a human, catches risky uses without ever having a person read the conversation itself. Think of airport security screening: instead of a person opening every bag to inspect what's inside, an X-ray machine flags whether something dangerous is present. OpenAI adopted this approach because it promised enterprise customers that it would not store the content of their API requests on its servers at all. If no data is retained, catching misuse has to be handled by an automated system rather than a human reviewer.

Once this step runs, if risk is detected, alerts go out to two places: the enterprise customer using the API, and OpenAI itself. But the alert OpenAI receives does not contain the actual conversation content—only the type and severity of the risk. Whether to actually disclose the content is left up to the customer. In other words, Automated Safety Review only handles detecting risk; the judgment call afterward is left to humans.

How it shows up in the news

The article explains that Automated Safety Review determines misuse without human review or OpenAI storing content on its servers. A common misunderstanding is that this doesn't mean the content is never looked at—the automated system does scan the content, but the results aren't stored in a human-readable form or passed directly to OpenAI staff.

See also

Stories using this term

Browse every entry