Skip to content

Topic

AI inference hardware

Purpose-built accelerators and rack-scale systems targeted at inference rather than training workloads.

Current stories

invest6 publishers

Volantis raises $88 million to link one GPU to 220 memory chips with light

Volantis, a San Francisco chip startup, closed an $88 million Series A to replace the electrical links between AI chips and memory with optical ones. Its A-1 system, promised to customers in 2027, puts a date on the bet that memory, more than compute, now limits AI inference.

Perspective Coverage

6 publishers
Builder
Builder 41%
Operator
Operator 16%
Investor
Investor 43%

Reality

Evidence40
Adoption5
Hype gap+35
Incentives60
Confidence60
build2 publishers

Cognition reports 4.8x token throughput on Nvidia Vera Rubin in its own coding-agent test

Cognition says its SWE-2 model produced up to 4.8 times more total token throughput on Nvidia's Vera Rubin NVL72 than on GB200, in a test on CoreWeave. The company-run result points to more agent capacity per system, but whether long coding tasks get cheaper depends on what the new systems cost per hour.

Reality

Evidence40
Adoption30
Hype gap+35
Incentives85
Confidence60
invest8 publishers

Marvell just put Google's purchase orders on its cap table

A warrant for 58.97 million shares vests one $500 million tranche of custom-chip revenue at a time, turning a supply agreement into an equity schedule that runs to 2033.

Perspective Coverage

8 publishers
Builder
Builder 14%
Operator
Operator 17%
Investor
Investor 69%

Reality

Evidence82
Adoption30
Hype gap+30
Incentives55
Confidence74
invest4 publishers

Positron's second tranche pays a 16 per cent step-up inside the same announcement

The $375m Series C priced Positron at $3.5bn before the money, and the up-to-$500m follow-on has to be sitting at $4.5bn pre-money to reach the $5bn headline, which only holds if every dollar of it is drawn.

Perspective Coverage

4 publishers
Builder
Builder 34%
Operator
Operator 21%
Investor
Investor 45%

Reality

Evidence55
Adoption30
Hype gap+35
Incentives70
Confidence60
product7 publishers

Apple weighs wiring an M8 Ultra AI server together with Nvidia's NVLink Fusion

The Information reports two configurations, one with two chips and one with four, aimed at developers, businesses and governments, with a possible 2029 launch. Apple already builds servers for Private Cloud Compute and has turned partners away from them.

Perspective Coverage

7 publishers
Builder
Builder 32%
Operator
Operator 32%
Investor
Investor 36%

Reality

Evidence35
Adoption15
Hype gap+20
Incentives55
Confidence40
build3 publishers

Apple's reported M8 Ultra server would pick up NVLink where UltraFusion stops

The Information reports Apple is weighing Nvidia's NVLink Fusion for an inference server built on two or four M8 Ultra chips in 2029. The engineering question is which part of that package would join the NVLink domain.

Perspective Coverage

3 publishers
Builder
Builder 37%
Operator
Operator 23%
Investor
Investor 40%

Reality

Evidence30
Adoption18
Hype gap+42
Incentives55
Confidence50

Earlier coverage

  1. Samsung puts MAC trees in every LPDDR5X bank because HBM costs too much

    Build · August 25, 2026 · 1 publisher

  2. Etched's $10.3B mark prices a non-Nvidia inference bet at ten times booked orders

    Invest · August 20, 2026 · 1 publisher