product1 distinct publisher
Kog's pitch: the cheapest inference upgrade is the H200s you already bought
The French startup's only public number is 3,000 tokens per second on a 2-billion-parameter model. Its 30x claim for real LLMs has not been shown yet.
Publishers:techcrunch.com
Reality
- Evidence30
- Adoption18
- Hype gap+52
- Incentives78
- Confidence38