Google says frontier models already know the facts they get wrong. That is a budget decision.
Google Research reports Gemini-3-Pro and GPT-5 encode 95-98% of tested facts yet fail to recall 26-34% of them, moving the fix from pretraining scale toward post-training and inference.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+20
- Incentives68
- Confidence48