Skip to content

Topic

Agentic Coding Models

AI language models designed to autonomously write, debug, and complete software engineering tasks, benchmarked on agentic coding performance and cost.

Current stories

buildOne report1 publisher

Failed runs decide whether Claude Sonnet 5.5 undercuts Opus 5.5

Anthropic prices Claude Sonnet 5.5 at half Opus 5.5's per-token rate, but Artificial Analysis measured $7.67 per task at max effort against Opus's $5.98. In The New Stack's own tests, runs that were billed but never finished decided which model cost less per completed job.

Publishers:thenewstack.io

Reality

Evidence55
Adoption
Insufficient
Hype gap+20
Incentives35
Confidence50
buildOne report1 publisher

Anthropic prices its newer Sonnet a third below Sonnet 4.5

Sonnet 4.5 still leads GPT-5 on the coding leaderboards, and GPT-5 lists about 46 percent below it on a 5:1 token mix. Anthropic's current Sonnet undercuts both of Sonnet 4.5's list prices, and that complicates a routing plan built on the older pair.

Publishers:dev.to

Reality

Evidence40
Adoption20
Hype gap+15
Incentives55
Confidence45