Skip to content

Build1 publisher2 min readPublished

Pichai ranked frontier AGI above Google Cloud in Alphabet's compute queue

Alphabet says it is supply-constrained at the component level, and Sundar Pichai has told analysts that frontier AGI work gets the compute first. A dev.to account reports Google refused Meta the Gemini capacity it asked for.

The Engineer · Build desk

Illustration accompanying Pichai ranked frontier AGI above Google Cloud in Alphabet's compute queue

What happened

  • Alphabet CFO Anat Ashkenazi has described the company as operating in a "supply-constrained environment", which the dev.to account treats as a real ceiling on accelerator supply.
  • DeepMind CEO Demis Hassabis has traced the bottleneck to a handful of component suppliers sitting behind every advanced accelerator on the market, not to anything Google has mismanaged.
  • Sundar Pichai told analysts how the capacity is ranked, putting frontier AGI work first as the foundation the rest of the company depends on and grouping Cloud with Search and YouTube for the remainder.
  • Around March 2026, according to the dev.to account, Google told Meta it could not supply the Gemini compute capacity Meta had requested, and Meta instructed staff to conserve their AI usage.
  • The same account points to a $920-million-a-month lease from an unnamed rocket company as the visible downstream trace of how tightly the constraint binds.

Compiled by The EngineerSomething wrong?How this is made

Why it matters

  • constraint Publishing a priority order leaves total demand where it was. The deprioritised half goes looking for capacity on some other vendor's schedule, at whatever price that vendor is asking.
  • exposure The largest draw on a constrained supplier is also the most legible line item to cut, and Meta was rationed because its demand was big enough to matter.
  • decision Buyers of accelerator capacity from a vendor that also runs its own frontier programme have to get the priority question into the contract, since the only public answer today is a transcript.
  • precedent A named executive stating the order out loud gives internal platform teams a defensible model for rationing GPUs, where the usual failure is that nobody has standing to refuse a request.

An allocation hierarchy decides which requests wait. Alphabet designs the TPU, builds it, and owns the fleet, and according to the dev.to account it still has to run an explicit prioritisation policy over that resource [12]. The binding input sits upstream of the design: high-bandwidth memory from a small number of manufacturers [3].

Downstream, the cost showed up as slipped internal work. The shortfall was large enough to disrupt several of Meta's internal AI projects, per the same account [6].

Multiply the lease's monthly figure by twelve and it runs to roughly $11.04bn a year [14]. The dev.to post refers to the lessor only as a rocket company and does not name it [15]. So the figure is a monthly rate as one publisher reported it.

Whether any of this transfers to your own cluster depends on which shortage you have. The test is whether deleting every unvalidated reservation from the queue would make the total fit; if it would, the fix is validation and a governance policy just formalises guesses. The dev.to author is explicit that Alphabet's case is not that one: its available capacity, its internal demand and its Cloud customers' demand are all real, with no dependence on a forecast that might later prove phantom [10]. The framing offered is what an organisation does when every signal is trustworthy and the total still does not fit [11].

One publisher carries the Meta episode, and the post is an argument. Alphabet's own executives carry the ceiling and the ordering, on the record, to analysts. If you buy capacity from a company that also runs a frontier lab, the question to settle before signing is what your contract says when the lab wants your slot.

What to watch

  • Whether Alphabet's next earnings call keeps frontier AGI ahead of Cloud in the stated order, or moves Cloud up.
  • Whether high-bandwidth memory output from that small group of manufacturers loosens enough to lift the ceiling.
  • Whether another Google Cloud customer says it was refused TPU capacity, which would make the Meta refusal a standing policy.
Loading claim ledger
Loading source directory links
Loading share composer
Loading topic controls
Loading related stories