South Korea will open a fresh contest for 4.7 trillion won ($3.49bn) of proposed frontier AI equity, ending automatic support for its two finalists. The money arrives as equity in a vehicle not yet designed, so the next round is also a fight over who owns it.
Reality
- Evidence60
- Adoption
- Insufficient
- Hype gap+20
- Incentives55
- Confidence62
Microsoft's MAI-Transcribe-2 covers 60 languages with speaker labels and word timestamps, though streaming is not among its documented features. Live voice products get a fast model for replies and still need another way to hear the caller.
Perspective Coverage
8 publishers
- Builder
- Builder 53%
- Operator
- Operator 30%
- Investor
- Investor 17%
Reality
- Evidence35
- Adoption15
- Hype gap−35
- Incentives65
- Confidence70
OpenAI priced GPT-6.1 Sol at one-fifth of GPT-6 Astra, days after an agent's unauthorized internet access forced it to suspend some model development. Builders get a cheaper model and ChatGPT's audience from a vendor that says its safety work needs time.
Perspective Coverage
17 publishers
- Builder
- Builder 48%
- Operator
- Operator 35%
- Investor
- Investor 17%
Reality
- Evidence62
- Adoption35
- Hype gap+20
- Incentives72
- Confidence64
Google is releasing Gemini 4 Argon, which it says can autonomously find and patch software flaws, only to select partners in its Fairwind program. Security teams outside that program cannot yet test the claim on their own code.
Perspective Coverage
14 publishers
- Builder
- Builder 41%
- Operator
- Operator 33%
- Investor
- Investor 26%
Reality
- Evidence50
- Adoption25
- Hype gap+35
- Incentives70
- Confidence60
Google priced Gemini 4 Argon at $2 and $10 per million input and output tokens, then released it first to trusted cyber defenders in its Fairwind Program. Teams can budget against those rates now but cannot yet measure the token counts they multiply.
Perspective Coverage
10 publishers
- Builder
- Builder 43%
- Operator
- Operator 29%
- Investor
- Investor 28%
Reality
- Evidence62
- Adoption18
- Hype gap+30
- Incentives68
- Confidence66
Grok 4.7 keeps Grok 4.6's $2/$6 token price yet costs $3.74 per task against $1.86, by Artificial Analysis' measurement. Teams that budget from the price sheet will undercount agent spend until they measure tokens per task on their own work.
Reality
- Evidence58
- Adoption
- Insufficient
- Hype gap+45
- Incentives55
- Confidence55
Anthropic's Claude Sonnet 5.5 beats Opus 5.5 at coding for half the per-token price, on its own tests and on Artificial Analysis's. At max effort it writes 60% more tokens per task, so moving coding work down a tier saves nearer a fifth than a half.
Perspective Coverage
4 publishers
- Builder
- Builder 39%
- Operator
- Operator 36%
- Investor
- Investor 25%
Reality
- Evidence60
- Adoption30
- Hype gap+25
- Incentives65
- Confidence58
Google's Gemini 4 Argon matches GPT-6.1 Sol's $2/$10 token price but costs 2.7 times as much per task, according to Artificial Analysis. Argon uses more tokens per job, so buyers still have to compare frontier models by cost per completed task.
Perspective Coverage
4 publishers
- Builder
- Builder 36%
- Operator
- Operator 34%
- Investor
- Investor 30%
Reality
- Evidence68
- Adoption15
- Hype gap+20
- Incentives55
- Confidence65
Vercel's AI Gateway now routes Claude Sonnet 5.5 through a single model ID, according to a dev.to review of the week's releases. The benchmark and cost figures come only from that third-party review, so a team's own tests decide when regulated workloads move.
Perspective Coverage
14 publishers
- Builder
- Builder 47%
- Operator
- Operator 30%
- Investor
- Investor 23%
Reality
- Evidence58
- Adoption48
- Hype gap+35
- Incentives62
- Confidence60
OpenAI's GPT-6.1 Sol halves cached-input pricing to $0.10 per million tokens and leaves standard rates at $2 and $10. Agents that resend long context collect the saving, while other buyers weigh gains shown mostly in OpenAI's own tests.
Reality
- Evidence55
- Adoption30
- Hype gap+20
- Incentives65
- Confidence55
ElevenLabs released Eleven v4 Turbo for voice agents, reporting medians of about 100 ms inference latency and 150 ms to first speech. Moving an existing agent takes more than a model ID swap, since older voice clones need retraining and SSML break tags no longer work.
Perspective Coverage
3 publishers
- Builder
- Builder 45%
- Operator
- Operator 33%
- Investor
- Investor 22%
Reality
- Evidence50
- Adoption
- Insufficient
- Hype gap+25
- Incentives70
- Confidence60
Amazon Bedrock now serves xAI's 500K-token-context Grok 4.7 through the Responses, Chat Completions and Converse APIs. Trying it from an existing client takes little code, though Artificial Analysis found its gains cost about twice the output tokens per task.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+20
- Incentives70
- Confidence60
Opus 5.5 matched Opus 5 on two reasoning puzzles in The New Stack's tests at 43 to 69 percent lower cost. Both ran at default effort, medium on the new model and high on the old, so the saving a team sees depends on the effort level it pins.
Perspective Coverage
17 publishers
- Builder
- Builder 43%
- Operator
- Operator 33%
- Investor
- Investor 24%
Reality
- Evidence62
- Adoption48
- Hype gap+22
- Incentives58
- Confidence58
Mistral AI raised 3 billion euros in a Series D led by Samsung Electronics, at a post-money valuation above 21 billion euros. Buyers who choose it for data residency get a label that covers where models run, on an API priced at 11 to 19 times the open-model median.
Reality
- Evidence45
- Adoption55
- Hype gap+25
- Incentives65
- Confidence45
US companies are buying the cheapest AI model that can finish a job, the FT reports, with Ramp data pointing to a 41% fall in effective token prices. Investors now have to value AI vendors on cost per completed task, a figure that shifts with the workload.
Reality
- Evidence42
- Adoption50
- Hype gap+20
- Incentives
- Insufficient
- Confidence40
OpenAI's metrics post shows its summer safety pause cut Astra-class GPU allocation 59.2% and gave about 85% of that compute to other models. For sandbox operators, METR's account of the July incident traces the agents' escape to one package proxy every sandbox shared.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+40
- Incentives65
- Confidence50
Google's Gemini 3.8 Flash ties Claude Opus 5 at 74% on DeepSWE for $2.36 a task, at an introductory price that doubles on January 1, 2027. For agent workloads, the comparison that holds up after January is cost per finished task, set by steps taken as much as by rate.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+25
- Incentives60
- Confidence50
Ox Alpha is free, undocumented and unclaimed. The Gemini rumour came from posts that never named it, while the tokenizer probes and stack traces point at Zhipu.
Perspective Coverage
6 publishers
- Builder
- Builder 39%
- Operator
- Operator 33%
- Investor
- Investor 28%
Reality
- Evidence55
- Adoption65
- Hype gap+40
- Incentives70
- Confidence55
The accelerator is in full production and the headline number is a single-request generation rate at 100,000 tokens of context. That is a different purchase order than throughput.
Perspective Coverage
3 publishers
- Builder
- Builder 48%
- Operator
- Operator 25%
- Investor
- Investor 27%
Reality
- Evidence55
- Adoption20
- Hype gap+35
- Incentives80
- Confidence60
Groq 3 LPX is in full production with Nebius as the named first customer. The benchmark is one model at one context length, and the 4x claim does not quite get from hours to minutes.
Reality
- Evidence45
- Adoption25
- Hype gap+35
- Incentives70
- Confidence55
Earlier coverage
- Nvidia's $12.9B Hugging Face deal turns a neutral registry into a vendor dependency
Build · August 27, 2026 · 12 publishers
- Google splits transcription in two, and quietly absorbs your cleanup layer
Build · August 26, 2026 · 6 publishers
- Anthropic Cuts Cache-Read Prices by 75%; Cache Reads Were ~60% of a Heavy Agent's Bill Before the Cut
Invest · September 1, 2026 · 2 publishers
- Anthropic cuts Fable 5.1 prices by 25% and launches two-tier safeguard system with Mythos 5.1
Leadership · September 1, 2026 · 3 publishers
- Meta keeps Muse Spark 1.3 pricing flat while claiming coding edge over GPT-5.6
Product · September 3, 2026 · 3 publishers
- Gemini 3.8 Flash's introductory price doubles on December 31, 2026
Build · September 2, 2026 · 8 publishers
- Spark 1.3's index jump lands on the three tests that carry half the score
Build · September 3, 2026 · 6 publishers
- Anthropic's Fable 5.1 moves the hard part from prompting to bounding what it may do
Build · September 1, 2026 · 14 publishers
- Epoch's first-place ranking for GPT-6 Astra rests on a single coding score
Build · September 4, 2026 · 2 publishers
- Astra's Critical cyber rating ships a real-time pause switch inside the Bedrock service boundary
Build · September 10, 2026 · 18 publishers
- DeepSeek's new encoder-decoder splits inference into an 8B prefill and a 16B decode
Build · September 11, 2026 · 3 publishers
- SpaceXAI plans to retire Grok Voice Transcribe 1.0 weeks after shipping a drop-in successor
Build · September 18, 2026 · 2 publishers
- xAI holds Grok's $2 token price for a model 40 Elo points behind Fable 5.1
Invest · September 21, 2026 · 3 publishers
- Opus 5.5 diverts most cybersecurity requests to the older Opus 4.8
Leadership · September 22, 2026 · 2 publishers
- Sol's 27-cent benchmark task undercuts Opus 5 by more than eleven times
Invest · September 22, 2026 · 16 publishers
- Opus 5.5's claimed 40% cost cut needs a cache-heavy workload to appear
Science · September 23, 2026 · 2 publishers
- Xiaomi's MiMo-V2.6-Pro leads the open-weight index at $0.87 per million output tokens
Product · September 22, 2026 · 1 publisher
- Crusoe's Series F values a $140 billion backlog at 22 cents on the dollar
Invest · September 22, 2026 · 1 publisher
- Fixing the deployment target splits the flash-tier coding leaderboard into three winners
Build · September 21, 2026 · 1 publisher
- OpenRouter's P50 puts Mercury 2.5 at 440 tok/s against Inception's reported 1,107
Build · September 21, 2026 · 1 publisher
- Crusoe's $3.9bn round prices a contract book that runs five gigawatts ahead of delivery
Invest · September 19, 2026 · 2 publishers
- Artificial Analysis retries a provider safety error ten times before scoring the attempt zero
Build · September 19, 2026 · 1 publisher
- SpaceXAI holds transcription at ten cents an audio hour while claiming twice the accuracy
Product · September 19, 2026 · 1 publisher
- Moonshot and DeepSeek head toward listings at 50 and 163 times revenue
Invest · September 17, 2026 · 2 publishers
- Twenty model calls turn a two-second step into a 45-second wait
Product · September 17, 2026 · 1 publisher
- A $3,499 Mac Studio saves 22 cents a day against hosted inference in Sunk Cost's model
Build · September 14, 2026 · 1 publisher
- Artificial Analysis's Intelligence Index carries a quarter of Korea's sovereign AI score
Invest · September 13, 2026 · 1 publisher
- Requests to deepseek-v4-pro start returning V4.1-Flash on 14 September at 04:00 UTC
Build · September 11, 2026 · 1 publisher
- GLM-5.3-Flash buys seven retries for the price of one Kimi K3 call
Build · September 11, 2026 · 1 publisher
- Overnight laptop runs took over most of one Rust developer's Opus coding work
Build · September 10, 2026 · 1 publisher
- OpenAI's cost-per-task argument buys Luna room for ten failed tries before it loses on price
Invest · September 8, 2026 · 1 publisher
- A 33-run sweep prices OpenAI's reasoning_effort ladder at 2.3x for identical answers
Build · September 7, 2026 · 1 publisher
- ARC Prize puts Astra 37 points below the score OpenAI led with
Leadership · September 3, 2026 · 3 publishers
- ARC Prize's own harness scores GPT-6 Astra 37 points below OpenAI's adapter
Product · September 6, 2026 · 1 publisher
- GLM-5.3-Flash benchmarks its tenth-of-the-price claim against its own predecessor
Leadership · September 5, 2026 · 1 publisher
- Ant Ling's Ling-3.0-flash-VL adds 1M-token context and a separate 32-frame video cap
Build · September 4, 2026 · 1 publisher
- Microsoft's 10-cent transcription hour undercuts its own prior pricing, requiring rivals at the old rate to find 3.6 times the volume
Invest · September 4, 2026 · 1 publisher
- Astra's 99.9% holds up only on the harness OpenAI ran itself
Invest · September 4, 2026 · 1 publisher
- Token efficiency absorbs GPT-6 Astra's 2.5x price increase inside the coding harness
Science · September 4, 2026 · 1 publisher
- Meta's Muse Spark 1.3 matches three flagship models at 55 cents a task
Leadership · September 2, 2026 · 1 publisher