Skip to content

Leadership1 publisher3 min readPublished

Frontier lab scientists ask governments to put auditors inside their own companies

Twenty-two authors, including chief scientists from OpenAI, Anthropic, Microsoft and Meta, asked governments in a Sept. 28 paper to require reporting on how far labs have automated their research and to embed independent auditors inside some firms.

The Board Room · Leadership desk

Illustration accompanying Frontier lab scientists ask governments to put auditors inside their own companies

What happened

  • Proposed tools include reporting of automation levels, auditors working inside some firms on the banking and nuclear regulatory model, a speed limit on capability growth, and air-gapped research networks.
  • Signatories include OpenAI's Jakub Pachocki, Anthropic's Jack Clark, Microsoft's Eric Horvitz and Meta's Dawn Song, alongside Turing Award winners Hinton, Bengio and Barto.
  • OpenAI, Microsoft and Meta declined to comment on the paper, and Anthropic did not respond.

Compiled by The Board RoomSomething wrong?How this is made

Why it matters

  • decision By naming the tools they want, from required reporting to embedded auditors, the authors are trying to shape the oversight regime before governments design one.
  • constraint Oversight rests on self-reported metrics like Anthropic's 26 percent; without a defined measurement standard, an outside auditor cannot verify how far automation has gone.
  • contradiction Insiders warn of extreme risk while Nvidia's Huang calls the warnings odd and a former CAISI head says embedded evaluators is ambiguous, so the proposal is contested even among practitioners.
  • precedent OpenAI's own goal of an automated AI researcher by March 2028 sets a timeline against which any embedded-auditor scheme would have to be built.

The people making this request write for the same companies that oversight would inspect. Jakub Pachocki is OpenAI's chief scientist, Jack Clark co-founded Anthropic, Eric Horvitz is Microsoft's chief scientific officer, and Dawn Song is Meta's vice president of AI research [5]. The paper says the authors wrote in a personal capacity and that their views do not necessarily represent their employers [6]. Academic and civil-society researchers initiated and led the project [7]. That distinction matters when the ask is for governments to look inside the firms these authors work for.

The measurement request is specific. Frontier companies would report how much of their research AI has automated, how fast capability is advancing, where research money goes, and whether AI systems take part in high-stakes research decisions [8]. Independent auditors could work inside some companies, on the model used by banking and nuclear regulators [9]. Song put the case for automated oversight plainly: "Already today, we are at the stage where we need AI systems to monitor what agents are doing. There is no other way to even observe and monitor these agents, humans are already insufficient," she said [10].

The one hard number in the paper is Anthropic's own. The authors cite the company's figure that AI completed 26 percent of its internal research and development work with only high-level supervision in August 2026, up from 1 percent in March [3]. Those figures are self-reported, and the authors say the threshold for an intelligence explosion has not been reached [4]. So the case for oversight rests partly on a metric that only the lab can produce. That is one reason the authors want required reporting: an outsider cannot check the 26 percent without a defined way to measure it.

The paper's central worry is a feedback loop. AI systems that help build more capable successors, which then do more of the research, is what the authors call an intelligence explosion [14]. Only thousands of people now do frontier AI research, and software agents can be copied and run in parallel, so full automation could add work equivalent to that of millions of researchers [15]. The paper's harder proposals follow from that: a speed limit on capability growth, ways to halt high-stakes experiments, and some automated research run on air-gapped networks physically isolated from the internet [11]. The authors say the extreme case could include "the marginalization or extinction of humanity" if highly capable systems escaped human control [12].

Not everyone reads the warnings the same way. Conrad Stosz, former head of CAISI, said it is "a little ambiguous what embedded evaluators means" [13]. Nvidia's Jensen Huang called the labs' warnings "odd" [16]. And the political ground is not neutral: President Trump told the United Nations last week that the United States would "totally reject" any "globalist scheme" to control AI [17]. OpenAI, Microsoft and Meta declined to comment on the paper, and Anthropic did not respond [18].

Pachocki has made a version of this argument from inside his own company. In a Sept. 6 essay he called for a "network of third-party auditors" to enforce common safety thresholds [19]. OpenAI aims to build an automated AI researcher by March 2028, a system that can set its own questions and run experiments with less human direction [20].

What to watch

  • Whether any government moves to require automation-level reporting or fund embedded auditors, and on what statutory basis.
  • Whether other frontier labs publish automation figures comparable to Anthropic's 26 percent, or a shared measurement standard emerges.
  • Whether OpenAI's automated AI researcher target of March 2028 holds, and what oversight is attached to it if it does.
Loading claim ledger
Loading source directory links
Loading share composer
Loading topic controls
Loading related stories