build1 distinct publisher
NVIDIA ships Groq 3 LPX and starts quoting inference in tokens per user, not per rack
The accelerator is in full production and the headline number is a single-request generation rate at 100,000 tokens of context. That is a different purchase order than throughput.
Publishers:nvidianews.nvidia.com
Reality
- Evidence26
- Adoption24