Blog vs Bytecode, a 28-item Kaggle benchmark, graded empty proxy responses as wrong and scored DeepSeek-R1 at 17% until a second gateway showed 100%. Once capture was fixed, frontier models lost points by flagging sound code, while a small Gemma model missed most of the planted flaws.
Reality
- Evidence40
- Adoption
- Insufficient
- Hype gap+10
- Incentives30
- Confidence45
The accelerator is in full production and the headline number is a single-request generation rate at 100,000 tokens of context. That is a different purchase order than throughput.
Perspective Coverage
3 publishers
- Builder
- Builder 48%
- Operator
- Operator 25%
- Investor
- Investor 27%
Reality
- Evidence55
- Adoption20
- Hype gap+35
- Incentives80
- Confidence60
Groq 3 LPX is in full production with Nebius as the named first customer. The benchmark is one model at one context length, and the 4x claim does not quite get from hours to minutes.
Reality
- Evidence45
- Adoption25
- Hype gap+35
- Incentives70
- Confidence55
NVIDIA's dense-versus-MoE explainer uses Nemotron 3.5 Lightning to walk through per-layer routing. Its own text puts the throughput number and the memory number on two different parameter counts.
Reality
- Evidence58
- Adoption20
- Hype gap+15
- Incentives82
- Confidence62
The first third-party benchmark of the LP30 rack came in at roughly four times the next-fastest public endpoint, measured one request at a time on a model small enough to fit.
Reality
- Evidence58
- Adoption20
- Hype gap+32
- Incentives74
- Confidence55
A single-box test in Japan put 76 tokens/s next to 4.4 tokens/s, then found the deciding variable elsewhere: which models held the output format and which invented reassurance.
Reality
- Evidence52
- Adoption22
- Hype gap−8
- Incentives45
- Confidence55