build1 distinct publisher
132 blockers, three defect families: the bigger model wrote better prose and the same bad plans
A 157-goal field test of an LLM plan-and-critique loop found failures clustered in three structural families. Swapping in gpt-4o changed the writing, not the dependency graph.
Publishers:dev.to
Reality
- Evidence44
- Adoption14
- Hype gap+16
- Incentives55
- Confidence37