AI GlossaryㅈSafety and controversy
Preparedness Framework
OpenAI's pre-set rulebook that automatically triggers safety measures once an AI model's dangerous capabilities reach a certain level
In plain words
The Preparedness Framework measures how close an AI model has come to dangerous capabilities using pre-defined thresholds, and it spells out in advance exactly what the company must do once those thresholds are crossed.
Think of it like triage in a hospital emergency room. If you sort patients into categories like low, medium, high, and critical, you can decide in advance what treatment applies at each stage. The Preparedness Framework works the same way. It breaks down a model's capabilities into levels across four areas—biological weapons, chemical weapons, cybersecurity, and AI self-improvement—and once a model reaches a certain level (especially the 'critical' level), a predetermined response kicks in automatically, such as halting a release or dramatically tightening security.
For example, the 'critical' level in the cybersecurity domain refers to a model being able to discover and exploit previously unknown security vulnerabilities in real, critical systems entirely on its own, without human help. If an assessment concludes that a model has reached this level, or that reaching it cannot be ruled out, enhanced security controls—like isolating access and encrypting data—follow immediately.
How it shows up in the news
Articles use it in phrasing like, "Astra was found to potentially meet the 'critical' cyber capability level under the Preparedness Framework." A common point of confusion is that this isn't a government-mandated legal regulation—it's an internal policy OpenAI created on its own in December 2023. That said, the White House has recently been moving to incorporate this kind of internal evaluation standard into government oversight frameworks, raising the possibility that a company's internal rules could migrate into public policy.
See also
Stories using this term
- OpenAI's Astra flagged for potential "Critical" cyber capabilityAI · 2026.08.09
- OpenAI's Upcoming Model Astra Flags Potential 'Critical' Cyber CapabilityAI · 2026.08.08
- OpenAI's upcoming model Astra approaches critical cyber capabilityAI · 2026.08.09
- OpenAI unveils GPT-6 Astra with better alignment and a critical risk ratingAI · 2026.09.04
- GPT-6 Astra's first 48 hours bring real-world use from architecture to roboticsThe Lab · 2026.09.06
- OpenAI's Astra Nears Launch Carrying 'Critical' Cyber RatingAI · 2026.09.02
