Skip to content

Science1 publisher2 min readPublished

Anthropic's reset provable-inference deadline lapses without a public update

Anthropic's Sept 30 deadline for Phase 1 of its provable-inference project, already pushed back 138 days, passed with no public update. The lapse shows how little of a dated pledge in its Responsible Scaling Policy anyone outside the company can check.

The Scientist · Science desk

Drafted by a language model from the sources cited here and checked against its claim ledger before publication. How we use AISend a correction

Illustration accompanying Anthropic's reset provable-inference deadline lapses without a public update
Generated illustration

What happened

  • Anthropic's original deadline was May 15, 2026, and on May 5 it moved the date to Sept 30, saying it needed to focus resources on a broader leveling-up initiative.
  • Provable inference is meant to sign model outputs so they trace back to specific weights, as a defence against attackers who alter a model after training.
  • Phase 1 is a planning and inventory milestone covering components, costs and timelines, and it does not deliver a working product.
  • The project builds on June 2025 Confidential Inference research into attestation using Trusted Platform Modules, which its authors called a sketch to start a conversation.
  • The FTC opened a probe into AI firms including Anthropic and OpenAI as the deadline arrived, an action Forkast calls the first US enforcement move on rogue AI agents.

Compiled by The ScientistSomething wrong?How this is made

Why it matters

  • contradiction Forkast treats Sept 30 as a live deadline, but the RSP page says the initial launch goals were replaced, so the date being watched may no longer be the one Anthropic is working to.
  • decision Anyone tracking RSP milestones as evidence of safety work has to read the policy page itself, because goals there can be replaced without a blog post or press release.
  • constraint Even a finished plan has to work around accelerators that do not fully support confidential computing, so hardware support limits how soon signed outputs can run in production.

Anthropic first gave itself 43 days for Phase 1, from April 2 to May 15 [1]. It moved the target ten days before that date came due [3]. The new window ran 181 days [4], a little over four times the first [5]. Forkast logs one more change to the record, a typo correction on July 29 [15].

For a milestone like this, the first question is what evidence of completion would look like. The deliverable is an inventory of components, costs and timelines [6]. Finished, it is a planning document. Unless Anthropic publishes it or says it exists, an outsider can test only the calendar. Forkast's test is the absence of a blog post, press release or social media update by close of business on Sept 30 [4].

That test is weaker than it looks. Anthropic's Responsible Scaling Policy page says the planned moonshot projects have launched and that the initial launch goals were replaced with more detailed objectives for ongoing work [9]. Forkast itself allows that silence does not equate to failure, and that it may reflect a change in internal reporting or a recalibrated scope [10]. The thing this doesn't tell you is whether the replacement objectives still hold a Phase 1 date, or what that date now is.

I think the lapse is weak evidence about the engineering and fair evidence about the policy's format. When the deliverable stays inside the company and the goals can be rewritten on the policy page, an outside reader cannot tell a missed deadline from a retired one.

The engineering constraint is older than the deadline. The June 2025 Confidential Inference authors noted that not all hardware accelerators fully support confidential computing environments [8]. Their approach tied trust to hardware-rooted attestation through Trusted Platform Modules [7], so signing outputs to specific weights needs chips that can produce that attestation. A Phase 1 plan of components and costs is where that gap would have to be costed.

Pawan Khandavilli, in an analysis Forkast cites, argued that the harder problem lies past inference: the industry, he wrote, lacks a framework for confidential agency, and agent identity today rests on trusting API keys and system prompts [12]. He listed four gaps: hardware-rooted agent identity, measured policy-as-code, portable attestation evidence and cross-cloud federation [13].

What to watch

  • Anthropic publishing the Phase 1 inventory of components, costs and timelines, or stating that the milestone was completed.
  • A dated provable-inference milestone appearing among the more detailed objectives on the Responsible Scaling Policy page.
  • Whether the FTC probe of Anthropic and OpenAI reaches the companies' internal safety-policy records.
Loading claim ledger
Loading source directory links
Loading share composer
Loading topic controls
Loading related stories