The accelerator is in full production and the headline number is a single-request generation rate at 100,000 tokens of context. That is a different purchase order than throughput.
Perspective Coverage
3 publishers
- Builder
- Builder 48%
- Operator
- Operator 25%
- Investor
- Investor 27%
Reality
- Evidence55
- Adoption20
- Hype gap+35
- Incentives80
- Confidence60
Groq 3 LPX is in full production with Nebius as the named first customer. The benchmark is one model at one context length, and the 4x claim does not quite get from hours to minutes.
Reality
- Evidence45
- Adoption25
- Hype gap+35
- Incentives70
- Confidence55
The 1.7x to 3.6x latency range is set by the baseline systems, not the chip, and the report's own publication date is unsettled. Read it as direction, not evidence.
Perspective Coverage
9 publishers
- Builder
- Builder 41%
- Operator
- Operator 31%
- Investor
- Investor 28%
Reality
- Evidence52
- Adoption14
- Hype gap+38
- Incentives82
- Confidence68
The $875 million round prices an inference chip whose performance figures come from cycle-accurate simulations and whose production date has moved from early 2027 into the second half. The system Positron ships today runs on HBM.
Reality
- Evidence42
- Adoption38
- Hype gap+45
- Incentives80
- Confidence52
The first third-party benchmark of the LP30 rack came in at roughly four times the next-fastest public endpoint, measured one request at a time on a model small enough to fit.
Reality
- Evidence58
- Adoption20
- Hype gap+32
- Incentives74
- Confidence55
d-Matrix says stacking compute on a co-designed DRAM die moves bits at roughly a sixth of HBM4's energy cost. It still has not said who fabricates the die.
Reality
- Evidence58
- Adoption12
- Hype gap+32
- Incentives72
- Confidence55
Intel's Hot Chips disclosure details 32 Xe Cores, 32MB of L2 and 16-deep matrix engines. The figure it does not include is the one that decides whether the capacity is servable.
Reality
- Evidence58
- Adoption
- Insufficient
- Hype gap+30
- Incentives68
- Confidence52
Talks over a partnership, investment or acquisition would give Nvidia an energy-efficient NPU line and the sovereign-AI customers already buying it, rather than leaving both to regional rivals.
Reality
- Evidence42
- Adoption38
- Hype gap+31
- Incentives66
- Confidence55
The first multi-wafer Cerebras system pairs a doubled clock with rebuilt power delivery and interconnect. The compute claims rest on WSE-3 Turbo dies that are otherwise unchanged.
Reality
- Evidence54
- Adoption34
- Hype gap+34
- Incentives76
- Confidence61
Etched has first-pass silicon, 400 staff and more than $1bn in booked orders. What it does not have in public is a peak FLOPS figure, a power draw, or a third-party benchmark.
Reality
- Evidence27
- Adoption41
- Hype gap+52
- Incentives79
- Confidence34
OpenAI's invite-only Ultrafast tier runs the same GPT-5.6 Sol up to 14 times quicker, while Google halves Gemini Flash pricing until December 31. Latency is now its own budget line.
Perspective Coverage
3 publishers
- Builder
- Builder 35%
- Operator
- Operator 33%
- Investor
- Investor 32%
Reality
- Evidence54
- Adoption42
- Hype gap+27
- Incentives74
- Confidence60