If agents can't do open-ended research, price compute against task automation
A new study says AI agents still fail at free-form research. The self-improvement story is the load-bearing beam under a lot of compute spending, and it just got a crack in it.
Reality
- Evidence22
- Adoption30
- Hype gap+38
- Incentives58
- Confidence