OpenAI is selling Dots, a cartoon agent that runs apps on a virtual machine, first to top-tier accounts like the $100-a-month Pro plan. Its errands for The Verge kept ending in human takeovers, so desk work in a team's own apps is the case to test.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+40
- Incentives45
- Confidence40
Manus 2.0 lets agents fire on email and Slack events and stay online on dedicated cloud computers, with a harness Manus says cost 32% less in one test. Manus has not described that test, so the 32% is a figure for its own workload until someone reruns it elsewhere.
Reality
- Evidence40
- Adoption
- Insufficient
- Hype gap+20
- Incentives60
- Confidence55
Instinct raised $1 billion at a $10 billion valuation, four times the $2.5 billion it was valued at a month earlier. Its agent runs errands by operating other companies' websites, apps and phone lines, so single-task apps end up serving software that acts for the user.
Perspective Coverage
3 publishers
- Builder
- Builder 27%
- Operator
- Operator 35%
- Investor
- Investor 38%
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+45
- Incentives60
- Confidence60
beans-picker cut a median macOS agent task from $0.388 to $0.060 in its author's benchmark by replacing full UI-tree reads with short candidate lists. Whether that carries over depends on how much of your own agent's bill goes to reading the window.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+20
- Incentives60
- Confidence40
The always-on agent product opened in beta on August 11. The docs cap accounts at 50 bots and chats, give every bot on an account the same computer, and skip Linux desktop entirely.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap+15
- Incentives40
- Confidence50
Generating the interface as video instead of rendering it from code is a serious research bet. Runway's own preview still lists legible text and long-session coherence as open problems, which is roughly where ordinary interfaces begin.
Perspective Coverage
5 publishers
- Builder
- Builder 47%
- Operator
- Operator 37%
- Investor
- Investor 16%
Reality
- Evidence45
- Adoption5
- Hype gap+35
- Incentives65
- Confidence60
Apple pulled a Mac mini and Mac Studio refresh forward to August 25 to meet buyers that, according to The Information, no engineering, developer relations or enterprise AI function inside the company was assigned to serve.
Perspective Coverage
3 publishers
- Builder
- Builder 25%
- Operator
- Operator 28%
- Investor
- Investor 47%
Reality
- Evidence35
- Adoption45
- Hype gap+30
- Incentives45
- Confidence40
Sensor Tower counted more than 900,000 downloads of Meta's Muse in its first week. Wired's reviewer found an agent that browses competently while its ideas tab keeps pitching a live link to the bank.
Reality
- Evidence45
- Adoption35
- Hype gap+30
- Incentives60
- Confidence50
Almost every announcement at the hour-long keynote tied back to Meta's personal agent, including glasses from $449 and a Mac that Muse can operate while the user is away. The metaverse barely came up.
Reality
- Evidence45
- Adoption20
- Hype gap+45
- Incentives72
- Confidence55
The company known for a pocket gadget now ships a cloud agent that installs on up to five Windows, macOS or Linux machines and runs on whatever model subscription the user already pays for. Four rival agents got there first.
Reality
- Evidence32
- Adoption12
- Hype gap+28
- Incentives72
- Confidence44
Sol lists at $2 and $10 per million tokens and Luna at $0.10 and $0.50. The per-task savings OpenAI published come mostly from the lower price, and the cheaper tier scores below its predecessor on computer use.
Publishers:forkast.news · openai.com Reality
- Evidence55
- Adoption32
- Hype gap+30
- Incentives82
- Confidence62
Cua's pitch is an OS-level driver that delivers clicks and keystrokes in the background on macOS, Windows and Linux. The post hedges it with "where supported by the platform". A team has to test that clause before designing around it.
Reality
- Evidence30
- Adoption
- Insufficient
- Hype gap+40
- Incentives
- Insufficient
- Confidence32
simframe's author rebuilt iOS Simulator perception as an always-on daemon after watching Claude Code take a fresh screenshot before every tap. The study he cites puts 75-94% of task time inside model calls.
Reality
- Evidence42
- Adoption10
- Hype gap+22
- Incentives55
- Confidence52
Cua published a 706,048-parameter model, its training data and its driver integration under MIT, so the claim that a small specialist can take over the per-field decision step is testable by anyone with forms.
Reality
- Evidence48
- Adoption12
- Hype gap+20
- Incentives68
- Confidence55
OpenAI's GPT-6 Astra video shows the model driving Blender on voice commands. The 3D and game-development communities spent the next two days telling OpenAI to stay away from human-made art, and a Fast Company column argues that better demos now land worse.
Reality
- Evidence34
- Adoption31
- Hype gap+22
- Incentives55
- Confidence41
Anthropic's support note says the mode picker goes away in a staged rollout, and accounts that move over cannot return to separate Chat and Cowork options. Connected apps stay live while a task runs.
Reality
- Evidence60
- Adoption25
- Hype gap+10
- Incentives70
- Confidence65
The new Junie CLI mode needs Docker, a generated build plan and a request specific enough to check, and the report it returns marks each check passed, failed or incomplete. JetBrains says it has run the agent on more than 1,500 of its own pull requests.
Reality
- Evidence45
- Adoption32
- Hype gap+18
- Incentives85
- Confidence55
Mininglamp published NavEval scores for its own model on its own benchmark. Across the three entries, the spread tracks how each stack reads a page. It is not a case of specialists beating frontier models.
Reality
- Evidence24
- Adoption12
- Hype gap+46
- Incentives86
- Confidence58
OpenAI paused new sign-ups to its $200 ChatGPT Pro plan because Astra demand outran its compute. It closed the tier to new buyers and left the price where it was, and existing Pro accounts keep running.
Reality
- Evidence58
- Adoption57
- Hype gap+24
- Incentives68
- Confidence62
Meta's AI unit wanted keystrokes and mouse movements from tens of thousands of colleagues, its chief technology officer told staff no one could opt out, and the program has been suspended since an internal leak this summer.
Reality
- Evidence66
- Adoption30
- Hype gap+20
- Incentives62
- Confidence64
Earlier coverage
- The Project button hands ChatGPT and Claude every file in the folder you point at
Product · September 8, 2026 · 1 publisher
- Nex-AGI's runnable N2.5 tiers ask for two H100s or sixteen H200s
Build · September 8, 2026 · 1 publisher
- Astra cuts the computer-use task from about 75 minutes to 40
Product · September 3, 2026 · 1 publisher
- CoArena licenses the preference labels its free leaderboard generates
Build · August 29, 2026 · 1 publisher
- One boolean on the click step decides whether the agent's browser launches at all
Build · August 29, 2026 · 1 publisher
- Anthropic ships a price dial with its new model, and that is now the buying decision
Leadership · August 26, 2026 · 1 publisher
- Computer use goes GA with a new request shape, and browser automation becomes something you buy
Build · August 25, 2026 · 1 publisher
- A 1MB script that never looks at the screen beats frontier models on computer-use benchmarks
Build · August 25, 2026 · 1 publisher
- Agents That Click: OpenAI Ships Computer Use, And Credential Policy Becomes Your Problem
Leadership · August 20, 2026 · 1 publisher
- Codex learns to click: the coding agent stops typing patches and starts operating the machine
Build · August 17, 2026 · 1 publisher