The study counts only complexity and dead code, because those are the two pyscn metrics that stay exact when you analyze just the files a patch touched. The human's own commit trips the same rule 24% of the time.
Reality
- Evidence55
- Adoption20
- Hype gap+12
- Incentives62
- Confidence60
Anthropic says multi-agent systems introduce their new problems in coordination, evaluation and reliability, and its own token arithmetic explains why picking a model is the cheaper half of the decision.
Reality
- Evidence45
- Adoption42
- Hype gap+15
- Incentives82
- Confidence55
Anthropic's 0.97 PGR belongs to small open-weights models, but the plumbing that produced it is copyable at roughly $22 per agent-hour, and the ten tests that judged the work still came from humans.
Reality
- Evidence45
- Adoption12
- Hype gap+20
- Incentives78
- Confidence55
A Samsung and University of Warsaw preprint tracks the names large language models keep inventing, Elena Vasquez and Marcus Chen among them, into hundreds of AI-generated papers that scholarly aggregators index without checking whether the author exists.
Reality
- Evidence58
- Adoption54
- Hype gap+14
- Incentives42
- Confidence56
Anthropic's temporary weekly headroom lapses at 11:59 PM PT with prices unchanged. Teams that wired CI pipelines and sprint plans to borrowed capacity now pay the same for roughly a third less work.
Reality
- Evidence34
- Adoption48
- Hype gap+18
- Incentives58
- Confidence40