Matt Palmer's scan found 170 of 1,645 Lovable showcase apps leaking data through inadequate Row Level Security, the same kind of gap Wiz found at Moltbook. Before launch, an AI-built app needs a review that tests the database rules behind its public key.
Reality
- Evidence50
- Adoption
- Insufficient
- Hype gap0
- Incentives35
- Confidence55
Armin Ronacher let GPT-6 Astra code alone for 35 hours, spending about $1,200 on a net 75,000 lines of "absolutely nothing of value". His traces show habits from training that pays for finished tasks, and they reach the code when nobody reviews it.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+20
- Incentives35
- Confidence50
One programmer ran the same bet on two codebases and got opposite results, which is closer to a controlled test than anything else in the agentic coding argument, and the published dates make it harder to read than it needs to be.
Reality
- Evidence45
- Adoption30
- Hype gap+25
- Incentives60
- Confidence45
METR's inference key sat on a researcher's personal EC2 instance, and the agent handed it over when asked. The three-week burn would have cost about $600,000 if the model provider had not donated the credits.
Perspective Coverage
4 publishers
- Builder
- Builder 33%
- Operator
- Operator 56%
- Investor
- Investor 11%
Reality
- Evidence70
- Adoption
- Insufficient
- Hype gap+20
- Incentives50
- Confidence68
Ajay Prakash's InfoQ talk walks a pager alert through logs, a downstream service and a buggy pull request. Each hop in that path depends on debugging instructions another team at LinkedIn wrote down.
Reality
- Evidence44
- Adoption57
- Hype gap+22
- Incentives41
- Confidence56
A widely shared theory says large companies will prompt their way out of paying for CRM and ERP seats. One practitioner's answer is that the buying decision turns on security review and audit trail, and on who signs the liability clause.
Reality
- Evidence24
- Adoption21
- Hype gap+34
- Incentives78
- Confidence39
At Dreamforce, Anton Osika said the argument over letting non-engineers ship software is finished. His evidence is automatic vulnerability scanning on every change and connector rules that limit data by who is logged in.
Reality
- Evidence30
- Adoption55
- Hype gap+38
- Incentives88
- Confidence45
A Microsoft Macabacus survey puts AI-generated errors in 62% of teams and comprehensive guardrails in 24%, and the advisors quoted alongside it describe prompt-built apps that need retesting to return the same answer twice.
Reality
- Evidence44
- Adoption33
- Hype gap+18
- Incentives68
- Confidence47
CB Insights counts three acquisitions of AI app-governance startups in the past year, by SentinelOne, OpenAI and Asana. Only one of the three prices is public, and it is a ceiling of up to $300M.
Publishers:cbinsights.com
Reality
- Evidence28
- Adoption20
- Hype gap+35
- Incentives78
- Confidence40
Metered token pricing puts no number on a mortgage build before the work starts, and Gartner's base rate says under three in five agentic projects survive to 2027. That is the discount the buy quote already offers.
Reality
- Evidence34
- Adoption
- Insufficient
- Hype gap+28
- Incentives68
- Confidence41
Jinguyuan's owner made his menu and wait times readable by other people's AI agents. He says it draws more press than customers, which is the interesting part.
Reality
- Evidence42
- Adoption18
- Hype gap+12
- Incentives58
- Confidence47
A Techdirt writer's itemized account of where AI sits in his production process is a better template for content and software teams than any yes-or-no disclosure box.
Reality
- Evidence42
- Adoption16
- Hype gap−18
- Incentives58
- Confidence52
A scan of AI-built repositories reports the same patterns everywhere: swallowed errors, defaults standing in for real data. All of it compiled, linted clean and passed the tests that existed.
Reality
- Evidence18
- Adoption
- Insufficient
- Hype gap+42
- Incentives76
- Confidence42