Manus's rebuilt 2.0 platform adds Cue, an app in which each personal agent has its own email, phone number, wallet and computer. Those agents can spend within set limits. A team that allows Cue at work has to set those limits first.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+30
- Incentives65
- Confidence45
GKE Agent Sandbox is generally available and publishes provisioning numbers a team can plan against. The scheduling and durability layers stacked on top of it rest on one public demo, and the deprecation argument lands there.
Reality
- Evidence48
- Adoption14
- Hype gap+58
- Incentives50
- Confidence55
An engineer writing on dev.to swapped the coordinating LLM in a multi-agent system for an XState machine and typed receipts, reporting a 70% token cut. Whether that transfers depends on what your coordinator cost.
Reality
- Evidence30
- Adoption15
- Hype gap+40
- Incentives45
- Confidence35
Business Insider has practitioners spending a fifth to a third of their working week supervising AI agents, work that looks like first-line management inside roles whose titles and pay bands have not moved with it.
Reality
- Evidence46
- Adoption41
- Hype gap+27
- Incentives63
- Confidence54
Tencent's Hunyuan Speech team put a swappable agent model behind a conversation model that re-decides every second whether to talk, and reports best-in-test timing on Full-Duplex-Bench v3 with a task-accuracy shortfall it describes only as slight.
Reality
- Evidence34
- Adoption12
- Hype gap+18
- Incentives62
- Confidence45
The desktop app creates a real git worktree for every task and launches a CLI agent inside it. The isolation is the same git command you already have. The product is the review loop built around it.
Reality
- Evidence42
- Adoption
- Insufficient
- Hype gap+14
- Incentives58
- Confidence48
Anthropic's redesigned Projects beta enforces one ceiling, 200 new threads a day, and it has not published a usage figure for a single thread. Subscribers see what a project drew only afterwards, in the Usage tab.
Reality
- Evidence66
- Adoption18
- Hype gap+14
- Incentives68
- Confidence61
shinpr has taken claude-code-workflows through 133 releases, and the recent ones delete structure the models no longer need. His session reader then found three mandatory steps in his own repository that never ran.
Reality
- Evidence42
- Adoption18
- Hype gap+12
- Incentives58
- Confidence46
The 14.13.1 release captures failed agent runs as incidents, promotes them into skill fixtures, and accepts an edit only when one split improves and neither regresses, all under an enforced dispatch budget.
Reality
- Evidence42
- Adoption15
- Hype gap+10
- Incentives65
- Confidence50
Factory raised $200 million to run language ports and migrations with orchestrated agents. Its projects open with a Readiness Report that checks whether the customer's repository is organized and documented well enough for those agents to work in.
Reality
- Evidence24
- Adoption33
- Hype gap+42
- Incentives72
- Confidence41
Kill an orchestrator and the fleet stops; kill a coordinator and every agent keeps typing. Foremerge's maintainer submitted the tool to a list of orchestrators anyway, because no list exists for the other layer.
Reality
- Evidence42
- Adoption
- Insufficient
- Hype gap+14
- Incentives78
- Confidence45
A dev.to post-mortem on Hermes puts the constraint plainly: workers never talk to workers. The promised comparison with LobeHub is missing from the text.
Reality
- Evidence20
- Adoption
- Insufficient
- Hype gap+35
- Incentives30
- Confidence45
James Coombs's pipeline now re-tiers any ticket touching three or more integration layers before an agent starts, a rule that came out of one ticket filed as an error handling improvement that ran 5 hours and 17 iterations.
Reality
- Evidence42
- Adoption18
- Hype gap+12
- Incentives35
- Confidence45
Anton Brilliantov spent eight parts specifying agent handoffs down to a single acceptance command. His ninth names the task shapes where those facts are still unknown, and the reading that has to happen first.
Reality
- Evidence32
- Adoption10
- Hype gap−10
- Incentives25
- Confidence45
Cursor's Projects beta gives each body of work a coordinator agent on its own cloud machine. Its internal deployment shows where the review load lands, and its productivity figures arrive without a baseline.
Reality
- Evidence44
- Adoption36
- Hype gap+33
- Incentives78
- Confidence57
Salesforce put Data 360, Informatica, MuleSoft, Tableau, Agentforce and Guardian under one new control plane on Thursday. The loops that run agents inside its own tools still belong to Mastra and the Claude Agent SDK.
Reality
- Evidence42
- Adoption15
- Hype gap+40
- Incentives75
- Confidence55
Lauren Tan's MIT-licensed plugin routes a described task into one of 23 playbooks, but the part that decides whether any of it transfers is the per-project skill that drives the real product and recognises failure.
Reality
- Evidence50
- Adoption30
- Hype gap+10
- Incentives40
- Confidence50
By Mistral's own account, agents handed one Fortran subroutine each returned working C++ that preserved the old design, and the multi-agent run stalled on hard bugs until a human operator unblocked it and reviewed the pull requests.
Reality
- Evidence42
- Adoption24
- Hype gap+20
- Incentives78
- Confidence55
A planning step asked for entities and got a bare array that JSON.parse waved straight through. The repair needed both a schema in the prompt and a wrapper in the parser, because two of the three backends behind the interface can only ask.
Reality
- Evidence45
- Adoption12
- Hype gap+12
- Incentives30
- Confidence52
Meta's terminal coding agent left beta on August 31 with plans from $5 to $50. That looks like a four-to-one discount on Codex and Claude Code seats, but line up the request bands and the gap narrows to two.
Reality
- Evidence57
- Adoption21
- Hype gap+27
- Incentives71
- Confidence63
Earlier coverage
- OpenAI's Assistants migration hands the tool loop back to your application code
Build · August 30, 2026 · 1 publisher
- A semaphore of ten turns a 50-call agent plan into five sequential waves
Build · August 27, 2026 · 1 publisher
- The agent bottleneck is a workflow engine your team wrote by accident
Product · August 26, 2026 · 1 publisher
- A mutex made of mkdir: the cheap fix for agent fleets fighting over one resource
Build · August 22, 2026 · 1 publisher
- Pocock's /wayfinder bets that the bottleneck in overnight agents is planning, not code
Build · August 20, 2026 · 1 publisher
- Invoked in three runs, executed in none: the cost rule that never got asked
Build · August 20, 2026 · 1 publisher
- 117 identical errors, zero bugs: when the defect lives in the orchestration
Build · August 18, 2026 · 1 publisher
- Cursor's undocumented 'desktop' command turns local AI agents into a scriptable control channel
Build · August 18, 2026 · 1 publisher
- The reason your agent gets worse after an hour is that nothing ever leaves the context window
Build · August 18, 2026 · 1 publisher