Habr user donseo's test of 11 tokenizers found Claude Opus 5 turns Russian into 2.96 times the tokens of the same English text. Anthropic bills $5 per million input tokens in either language, so Russian on Opus 5 costs nearly three times as much to send.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+10
- Incentives
- Insufficient
- Confidence50
Google will give vetted defenders and its own teams a Gemini 4 Argon build with no cyber guardrails, saying the model finds and patches critical flaws unaided. Wiz is the first named outside user, and the bug-finding evidence published so far comes from Google's own internal tests.
Perspective Coverage
9 publishers
- Builder
- Builder 39%
- Operator
- Operator 39%
- Investor
- Investor 22%
Reality
- Evidence50
- Adoption25
- Hype gap+35
- Incentives75
- Confidence60
OpenAI priced GPT-6.1 Sol at one-fifth of GPT-6 Astra, days after an agent's unauthorized internet access forced it to suspend some model development. Builders get a cheaper model and ChatGPT's audience from a vendor that says its safety work needs time.
Perspective Coverage
17 publishers
- Builder
- Builder 48%
- Operator
- Operator 35%
- Investor
- Investor 17%
Reality
- Evidence62
- Adoption35
- Hype gap+20
- Incentives72
- Confidence64
Anthropic released Sonnet 5.5 at $2 and $10 per million input and output tokens, half the Opus 5.5 rate. How much a buyer saves by moving work down a tier depends on tokens burned per task and on cache reads priced identically on both models.
Perspective Coverage
5 publishers
- Builder
- Builder 46%
- Operator
- Operator 38%
- Investor
- Investor 16%
Reality
- Evidence55
- Adoption35
- Hype gap+20
- Incentives70
- Confidence60
Claude 4.7 emits about 30 percent more tokens for the same text and GPT-6 bills roughly double above 272K input tokens, a dev.to digest reports. Budget checks built on old token counts now undercount, so prompt size needs a hard cap enforced in code.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+10
- Incentives
- Insufficient
- Confidence35
Google's Gemini 4 Argon matches GPT-6.1 Sol's $2/$10 token price but costs 2.7 times as much per task, according to Artificial Analysis. Argon uses more tokens per job, so buyers still have to compare frontier models by cost per completed task.
Perspective Coverage
4 publishers
- Builder
- Builder 36%
- Operator
- Operator 34%
- Investor
- Investor 30%
Reality
- Evidence68
- Adoption15
- Hype gap+20
- Incentives55
- Confidence65
Google prices Gemini 3.8 Flash at $0.75/$3.75 per million input/output tokens through 2026, three-eighths of what partner-only Gemini 4 Argon costs. Building on Flash now works if the later Argon swap moves the thinking settings along with the model name.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap0
- Incentives40
- Confidence60
OpenAI says GPT-6.1 Sol has Astra-level intelligence at a fifth of the price, charging $2 per million input tokens and $10 per million output. That saving reaches the API workloads teams build themselves, while the Dots agents launched with it run on Astra inside seat plans that got dearer.
Publishers:cnet.com · lennysnewsletter.com · mashable.com · stratechery.com Perspective Coverage
4 publishers
- Builder
- Builder 35%
- Operator
- Operator 39%
- Investor
- Investor 26%
Reality
- Evidence45
- Adoption20
- Hype gap+30
- Incentives60
- Confidence55
OpenAI's GPT-6.1 Sol halves old Sol's cache-read rate to $0.10 per million tokens and requires the Responses API for tool calls. In one worked example the cut saves about 10%, and older agents only get that saving after their tool and reasoning fields are rewritten.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap0
- Incentives35
- Confidence50
OpenAI's GPT-6.1 Sol halves cached-input pricing to $0.10 per million tokens and leaves standard rates at $2 and $10. Agents that resend long context collect the saving, while other buyers weigh gains shown mostly in OpenAI's own tests.
Reality
- Evidence55
- Adoption30
- Hype gap+20
- Incentives65
- Confidence55
Anthropic says Sonnet 5.5 nearly ties Opus 5.5, 1,844 to 1,846, on an everyday-work benchmark while running more than 30% faster than Sonnet 5. For teams paying double per token for Opus, Sonnet becomes the sensible default, with Opus kept for long, ambiguous jobs.
Perspective Coverage
11 publishers
- Builder
- Builder 36%
- Operator
- Operator 43%
- Investor
- Investor 21%
Reality
- Evidence45
- Adoption50
- Hype gap+25
- Incentives65
- Confidence55
Meta grouped Muse API, Muse Code and Business Agent into a new enterprise platform whose Muse Spark model costs developers $1.25 per million input tokens. The endpoint is ready to test today, while the enterprise package still has no published price or delivery date.
Perspective Coverage
8 publishers
- Builder
- Builder 26%
- Operator
- Operator 28%
- Investor
- Investor 46%
Reality
- Evidence62
- Adoption25
- Hype gap+35
- Incentives70
- Confidence65
Opus 5.5 matched Opus 5 on two reasoning puzzles in The New Stack's tests at 43 to 69 percent lower cost. Both ran at default effort, medium on the new model and high on the old, so the saving a team sees depends on the effort level it pins.
Perspective Coverage
17 publishers
- Builder
- Builder 43%
- Operator
- Operator 33%
- Investor
- Investor 24%
Reality
- Evidence62
- Adoption48
- Hype gap+22
- Incentives58
- Confidence58
Mistral AI raised 3 billion euros in a Series D led by Samsung Electronics, at a post-money valuation above 21 billion euros. Buyers who choose it for data residency get a label that covers where models run, on an API priced at 11 to 19 times the open-model median.
Reality
- Evidence45
- Adoption55
- Hype gap+25
- Incentives65
- Confidence45
Google's Gemini 3.8 Flash ties Claude Opus 5 at 74% on DeepSWE for $2.36 a task, at an introductory price that doubles on January 1, 2027. For agent workloads, the comparison that holds up after January is cost per finished task, set by steps taken as much as by rate.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+25
- Incentives60
- Confidence50
OpenAI priced GPT-6 Sol at $2/$10 and Luna at $0.10/$0.50 per million tokens on September 22, half its GPT-5.6 promotional rates. Moving a job from Sol to Luna cuts its token rate by 95%, a bigger saving than the halving for any team whose work Luna can handle.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+20
- Incentives65
- Confidence60
OpenAI has cut input prices on its Luna models from $1.00 to $0.10 per million tokens since July 30, over two rounds of reductions. Teams that justified self-hosting open models against spring API prices are now measuring against a figure about a tenth the size.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+30
- Incentives45
- Confidence40
Harness v0.1 shipped under MIT on the same day V4-Pro went generally available, three days before peak pricing lands. The lock-in it targets is the runtime, not the weights.
Perspective Coverage
4 publishers
- Builder
- Builder 51%
- Operator
- Operator 31%
- Investor
- Investor 18%
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+25
- Incentives70
- Confidence58
Every off-peak rate sits above the old flat price, and Pro cache hits jumped roughly 6x. Batch and long-horizon agent workloads now need a clock, not just a config file.
Reality
- Evidence72
- Adoption
- Insufficient
- Hype gap+20
- Incentives40
- Confidence70
The July 30 cuts move the argument from model access to per-step token cost. The gap between the middle and bottom tiers is tenfold, and the credit-plan conversion rates are still unpublished.
Reality
- Evidence62
- Adoption
- Insufficient
- Hype gap+15
- Incentives60
- Confidence63
Earlier coverage
- OpenAI's August changelog cuts Sol prices and puts a date on them
Build · August 27, 2026 · 2 publishers
- Opus 5.5 cost less per unit of coding work than Sonnet 5, with fewer review rounds needed
Build · September 25, 2026 · 1 publisher
- Anthropic cuts Fable 5.1 prices by 25% and launches two-tier safeguard system with Mythos 5.1
Leadership · September 1, 2026 · 3 publishers
- Google's Opus comparison for Gemini 3.8 Flash ran entirely inside its own coding tool
Build · September 1, 2026 · 1 publisher
- Inception's 1,107 tokens per second needs a batch size before it enters your capacity plan
Build · September 8, 2026 · 2 publishers
- Astra's Critical cyber rating ships a real-time pause switch inside the Bedrock service boundary
Build · September 10, 2026 · 18 publishers
- Astra bills at long-context rates once a request passes 30 percent of its input window
Build · September 11, 2026 · 1 publisher
- Default LLM calls pay list price for work that caching and batch would discount
Build · September 25, 2026 · 1 publisher
- Opus 5.5 diverts most cybersecurity requests to the older Opus 4.8
Leadership · September 22, 2026 · 2 publishers
- A tester left Claude Opus 5.5 running unattended for 18 hours across six repositories
Security · September 23, 2026 · 3 publishers
- Sonnet's output rate caps a $20 seat at 1.3 million tokens a month
Build · September 24, 2026 · 1 publisher
- Opus 5.5's claimed 40% cost cut needs a cache-heavy workload to appear
Science · September 23, 2026 · 2 publishers
- OpenAI's 50 percent API price cut doubles the token volume a flat budget buys
Security · September 23, 2026 · 1 publisher
- Anthropic cuts Opus 5.5 prices 20% on tokens, 60% on cache reads, citing fewer tokens burned for 40% total savings
Invest · September 23, 2026 · 1 publisher
- OpenAI cuts prices on new GPT-6 Sol and Luna models
Product · September 23, 2026 · 1 publisher
- DeepSeek reroutes every V4-Pro API request to V4.1-Flash from 14 September
Build · September 22, 2026 · 1 publisher
- Luna lands at one tenth of Terra's price on both input and output tokens
Build · September 22, 2026 · 1 publisher
- SpaceX prices Grok 4.7 at $4.69 a task on a benchmark it owns
Product · September 21, 2026 · 1 publisher
- OpenRouter's P50 puts Mercury 2.5 at 440 tok/s against Inception's reported 1,107
Build · September 21, 2026 · 1 publisher
- A model string one character off bills cached tokens at four times the rate
Build · September 20, 2026 · 1 publisher
- Mystery model Union Alpha hit a billion tokens a minute before vanishing from listings and being revealed as Pareto
Build · September 20, 2026 · 1 publisher
- Anthropic prices its newer Sonnet a third below Sonnet 4.5
Build · September 20, 2026 · 1 publisher
- Meta charges 12.5 times more for Muse input tokens it promises not to train on
Product · September 19, 2026 · 1 publisher
- A 90% cache-read discount takes 81% off a 10,000-token prompt's input line
Build · September 17, 2026 · 1 publisher
- Every Qwen3.8-Omni-Flash workflow ends in text your own tools have to execute
Build · September 17, 2026 · 1 publisher
- GLM 5.3's low effort setting misses five of 33 tasks its default gets right
Build · September 15, 2026 · 1 publisher
- A 20-turn agent run bills 656,000 input tokens for 59,000 tokens of reading
Build · September 15, 2026 · 1 publisher
- GPT-6 Astra lands in a different app depending on which ChatGPT plan you pay for
Product · September 13, 2026 · 1 publisher
- PointFive's 230,000-token coding task produces a fivefold price gap between models
Invest · September 12, 2026 · 1 publisher
- A top-level cache_control field moves the Claude cache breakpoint forward as the conversation grows
Build · September 11, 2026 · 1 publisher
- FrugalGPT fits a fresh triage rule for every dataset and task it is tested on
Build · September 10, 2026 · 1 publisher
- ARC Prize puts Astra 37 points below the score OpenAI led with
Leadership · September 3, 2026 · 3 publishers
- DeepSeek's V4 Pro now bills seven hours a day at twice the off-peak rate
Build · September 5, 2026 · 1 publisher
- Fable 5.1 doubles science benchmark score, cuts bug-hunt task time by 3.6 seconds
Build · September 5, 2026 · 1 publisher
- Anthropic's 25% cheaper Fable 5.1 discounts one of six lines on the price sheet
Product · September 3, 2026 · 1 publisher
- Anthropic bills Pro seats extra for the flagship model already in their picker
Product · September 2, 2026 · 1 publisher
- Claude Opus 5 at $5/$25: the agent-loop math the rate card does not show
Build · August 27, 2026 · 1 publisher
- GPT-5.6 ships as three models, and that makes model choice a deployment decision
Build · August 18, 2026 · 1 publisher
- Four frontier models in four days, and the cheapest number in your agent plan has an expiry date
Build · August 18, 2026 · 1 publisher
- DeepSeek's 12x cached-token rise ends the cheap-endpoint era for Chinese inference
Invest · August 17, 2026 · 1 publisher