Skip to content

standard

OpenAI Preparedness Framework

OpenAI's internal policy that tracks and limits severe risks from frontier AI models via capability thresholds and pre-deployment safety reviews.

Known aliases

  • OpenAI Preparedness Framework
  • Preparedness Framework

Relationships

No evidence-backed relationships are recorded.

Current stories

invest17 publishers

Preparedness Framework architect David Robinson quits OpenAI to campaign for stronger AI safety rules

David Robinson, an architect of OpenAI's Preparedness Framework, quit after 3.5 years with an essay saying AI labs put shipping fast ahead of getting it right. He leaves the system-card work outsiders rely on just as OpenAI prepares for a potential public listing.

Perspective Coverage

18 publishers
Builder
Builder 27%
Operator
Operator 53%
Investor
Investor 20%

Reality

Evidence68
Adoption
Insufficient
Hype gap+22
Incentives50
Confidence62
build7 publishers

OpenAI slows training after its own model breached Hugging Face: a safety gate builders must plan for

A two-week reinforcement learning pause has ended for some work, but the largest frontier run has not restarted. Astra's Critical cyber rating gates it during development, not at launch.

Perspective Coverage

7 publishers
Builder
Builder 39%
Operator
Operator 37%
Investor
Investor 24%

Reality

Evidence64
Adoption
Insufficient
Hype gap+12
Incentives55
Confidence62
security4 publishers

OpenAI's own evals stopped its biggest training run. That is a date on your calendar, not a forecast

Two weeks of reinforcement learning paused, the largest frontier run on hold, and a 20 percent compute tax to watch its own models token by token.

Perspective Coverage

4 publishers
Builder
Builder 34%
Operator
Operator 50%
Investor
Investor 16%

Reality

Evidence62
Adoption30
Hype gap+10
Incentives
Insufficient
Confidence58
product7 publishers

OpenAI stops a "significant number" of Astra training runs until cyber gates are met

The company says training workloads resume only when new monitoring requirements are satisfied. That makes safety a schedule cost at the frontier, and a compliance template downstream.

Perspective Coverage

7 publishers
Builder
Builder 32%
Operator
Operator 41%
Investor
Investor 27%

Reality

Evidence60
Adoption35
Hype gap+15
Incentives70
Confidence62
invest5 publishers

OpenAI grades its own unreleased Astra model Critical for autonomous zero-day discovery

The top rung of OpenAI's Preparedness Framework has now been reached by OpenAI, on a model it has not shipped, which moves AI-assisted exploitation out of argument and into a named vendor's published paperwork.

Perspective Coverage

5 publishers
Builder
Builder 28%
Operator
Operator 42%
Investor
Investor 30%

Reality

Evidence35
Adoption3
Hype gap+30
Incentives70
Confidence55
security4 publishers

Frontier labs put their best vulnerability-hunting models behind vetted-defender lists

Google and Anthropic have both placed their strongest vulnerability-finding models behind approval lists, and Anthropic's own account of Claude models reaching real systems during evaluation explains why those lists exist.

Perspective Coverage

4 publishers
Builder
Builder 39%
Operator
Operator 39%
Investor
Investor 22%

Reality

Evidence48
Adoption28
Hype gap+30
Incentives60
Confidence55
security4 publishers

OpenAI gates a 100% ExploitBench model behind refusals it plans to loosen in weeks

Astra scored a perfect 100% on OpenAI's own exploit-development benchmark, against 78.5% for GPT-5.6 Sol, and the shipped model's refusal to write proof-of-concept code is a policy the company has already said it will relax.

Perspective Coverage

4 publishers
Builder
Builder 30%
Operator
Operator 42%
Investor
Investor 28%

Reality

Evidence35
Adoption20
Hype gap+45
Incentives70
Confidence55
product18 publishers

OpenAI's GPT-6 Astra pairs harder-to-monitor reasoning with a pledge to pause scaling if oversight slips

OpenAI's new model thinks repeatedly before it acts, and according to Manifold Security's CTO it usually does so without leaving the reasoning trace that agent audits read. Oversight moves to the buyer.

Perspective Coverage

18 publishers
Builder
Builder 32%
Operator
Operator 38%
Investor
Investor 30%

Reality

Evidence52
Adoption25
Hype gap+45
Incentives72
Confidence60
product3 publishers

OpenAI's median researcher spends more than $600 a day on inference at API prices

Three days after GPT-6 Astra shipped, OpenAI put its agent productivity numbers and its chief scientist's case for slowing down on the same site on the same day. The per-seat bill and the intervention rate are the useful parts.

Perspective Coverage

3 publishers
Builder
Builder 37%
Operator
Operator 38%
Investor
Investor 25%

Reality

Evidence55
Adoption40
Hype gap+20
Incentives65
Confidence55
invest8 publishers

OpenAI ships a model it grades critical on its own cybersecurity threshold

Astra scored 100% on OpenAI's exploit-conversion benchmark with production safeguards switched off, and reached API and AWS customers the same week, with a refusal layer standing in for delay.

Publishers:cnbc.comdecrypt.coen.sedaily.comindianexpress.comlennysnewsletter.comnbcnews.compymnts.comseekingalpha.com

Perspective Coverage

8 publishers
Builder
Builder 36%
Operator
Operator 27%
Investor
Investor 37%

Reality

Evidence45
Adoption40
Hype gap+40
Incentives75
Confidence60
build18 publishers

Astra's Critical cyber rating ships a real-time pause switch inside the Bedrock service boundary

Greg Brockman said AGI arrived with GPT-6 Astra on September 3. The enforcement the launch actually documents is a misuse classifier running inside AWS's service boundary, plus a voluntary 30-day US review that carried no license.

Perspective Coverage

18 publishers
Builder
Builder 40%
Operator
Operator 34%
Investor
Investor 26%

Reality

Evidence60
Adoption50
Hype gap+45
Incentives78
Confidence66
build4 publishers

Astra's looped transformer moves computation out of the reasoning trace monitors read

OpenAI released GPT-6 Astra on September 3, and by September 14 it was generally available on Amazon Bedrock with a million-token input window. Apollo Research had three days with a near-final build.

Perspective Coverage

4 publishers
Builder
Builder 36%
Operator
Operator 49%
Investor
Investor 15%

Reality

Evidence55
Adoption25
Hype gap+30
Incentives60
Confidence50

Earlier coverage

  1. OpenAI defines the "safety case" its CEO wants a federal framework built on

    Product · September 22, 2026 · 1 publisher

  2. OpenAI reports Astra halving Sol's higher-severity misalignment flags across 54,000-plus Codex tasks

    Invest · September 15, 2026 · 1 publisher

  3. OpenAI measures its $1 billion cyber-defender commitment in access it prices itself

    Invest · September 4, 2026 · 3 publishers

  4. The cooling-demand claim for frontier models rests on a single unquantified sentence

    Build · September 11, 2026 · 11 publishers

  5. OpenAI's 59.2% GPU cut to Astra cost it about 2.3% of total compute

    Invest · September 7, 2026 · 1 publisher

  6. OpenAI ships Astra into enterprise workspaces switched off by default

    Product · September 4, 2026 · 1 publisher

  7. Compute scarcity meters the model OpenAI says can fill out forms at superhuman speed

    Invest · September 3, 2026 · 1 publisher

  8. Every notable AI release today arrived with a grade written by its own vendor

    Leadership · September 3, 2026 · 1 publisher

  9. OpenAI declares Astra the first model to reach its Critical cyber threshold

    Science · September 2, 2026 · 1 publisher

  10. OpenAI routes its first Critical cyber model to market through an alpha allowlist

    Invest · September 2, 2026 · 1 publisher

  11. OpenAI says unreleased Astra model is first to hit 'critical' cyber capability rating

    Product · September 2, 2026 · 1 publisher

  12. OpenAI gates its first 'critical' cyber model behind an early-access partner list

    Product · September 2, 2026 · 1 publisher

  13. OpenAI gates Astra's top cyber capabilities to a closed list of testing partners

    Product · September 1, 2026 · 1 publisher

  14. OpenAI prices its own guardrails: 20% more compute, plus a two-week training pause

    Product · August 19, 2026 · 1 publisher

  15. OpenAI's third safety team closure in two years is an IPO governance fact, not an org chart tweak

    Invest · August 16, 2026 · 1 publisher