build1 publisher
A 33-run sweep prices OpenAI's reasoning_effort ladder at 2.3x for identical answers
Eleven tasks with pre-computed answer keys, three runs each, seven effort settings. Everything from low upward scored 33 of 33, so the only thing the top rung buys is the number on the launch page.
Publishers:dev.to
Reality
- Evidence62
- Adoption45
- Hype gap+15
- Incentives55