Leadership1 distinct publisher3 min readUpdated
OpenAI is rolling browser and app control out to customers while Anthropic pushes Claude Cowork. The open question is not capability but which credentials the software gets, and who reviews the log.
The Board Room · Leadership desk
Compiled by The Board RoomSomething wrong?How this is made
OpenAI is rolling browser and app control out to customers while Anthropic pushes Claude Cowork. The open question is not capability but which credentials the software gets, and who reviews the log.
OpenAI has started rolling out its Computer Use tools to customers, and in the past year it has shipped a Chrome extension that lets ChatGPT take over the browser, cloud and in-app browsers for interacting with public websites, and Computer Use inside its Codex coding tool [1][2]. That is four separate surfaces where software now acts inside systems rather than answering questions about them [15], which makes credentials, scope, and audit an engineering decision rather than a procurement footnote.
Anthropic was first to market with its version in 2024 and is still working on it, and as of May at least 600,000 organizations had tried its Claude Cowork feature, which uses the tool, according to Business Insider [6][7]. OpenAI says its own capabilities have caught up [8], and employees who build the tools told Business Insider that there is room for improvement but that the tools are at an inflection point [5]. Read those two statements together before you plan a rollout: the people shipping it are saying it is good enough to hand work to and not finished.
The mechanics are the part worth reading twice. Previously, per Computer Use manager Ari Weinstein, ChatGPT took a screenshot, analyzed pixels, injected a command, and repeated; now, with a website open, it reads the page's memory structure, accessibility information, and link connections [9]. It still takes screenshots, and users are asked to let ChatGPT record their screen during setup so it can rapidly analyze what is on screen [10]. A tool that reads page internals and records the display is a data-handling question and an access question at the same time, and the two land on different owners in most organizations.
Unattended operation is already the selling point. Cristian Medina Ruiz, a hobbyist coder in the Czech Republic, used Codex with Computer Use to verify a rebuild of the 2013 game SimCity, and now leaves code running and iterating when he is away from the machine: "It does these things even when I'm away" [12]. Weinstein's framing is broader: "Once ChatGPT can use computers and software faster than you or I can, it's going to change the way that you, by default, want to interact with your computer" [13]. Default is the operative word. Defaults are set by whoever configures the extension first, which is usually not the security team.
Three decisions cannot be deferred. Which identity the agent uses, given that a browser takeover inherits whatever the operator is already signed into. Which surfaces are in scope, because OpenAI says the same underlying technology lets ChatGPT complete tasks in other apps, and president Greg Brockman wrote on X that an April update made the technology "no longer just for coders, but for anyone who does computer work" [3][4]. And what a reviewable record of agent actions looks like, since screenshots and page reads are inputs to the model, not an audit trail for you.
Watch whether that 600,000 trial figure converts into standing deployments [7], and whether either vendor ships admin-side scoping and logging as fast as it ships new surfaces. OpenAI declined to give specifics about its latest training data [14]; expect the same reticence about what the agent recorded on the way to finishing a task.
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
As of May, at least 600,000 organizations had tried Anthropic's Claude Cowork feature, which uses the Computer Use tool.
The tool still takes screenshots; users are asked to let ChatGPT record their screen as they set up Computer Use, so it can rapidly analyze what is on screen.
OpenAI declined to provide specifics about its latest Computer Use training data; Columbia University researcher Zhou Yu said the training likely included frame-by-frame screenshots at the base stage, human demonstration datasets, and reinforcement learning on virtual tasks.
OpenAI now has enough faith in the technology to start rolling out Computer Use tools to customers.
This year OpenAI launched a Chrome extension that lets ChatGPT take over the browser, created cloud and in-app browsers for ChatGPT to interact with public websites, and added Computer Use to its Codex coding tool.
The same underlying technology now also lets users have ChatGPT complete tasks on other apps.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Single-publisher access feature, mostly vendor-sourced
All material comes from one Business Insider article built on interviews with OpenAI's own Computer Use and browser staff. Capability and architecture claims are first-party and unverified; the one independent voice (Columbia's Zhou Yu) speaks inferentially about training and explicitly cautions about generalization. No benchmark result, no documentation, no second publisher, and OpenAI declined to detail training data.
Broad trial on the rival side, thin depth-of-use evidence
There is real shipped surface area (four OpenAI Computer Use entry points reaching customers) and one sizable third-party-facing number — at least 600,000 organizations having tried Claude Cowork. But 'tried' is trial breadth, not retention, and the only concrete user story is a single hobbyist coder. Task categories cited (data entry, compliance, scheduling) are staff observations without volumes.
Inflection-point framing outruns the article's own limitations
The narrative ('inflection point', capabilities 'caught up', changing how you want to use your computer by default) sits above what the same piece evidences: an independent researcher describing a persistent speed and generalization gap, admitted weakness on email and social feeds, no benchmark behind the parity claim, and a lone hobbyist as the illustrative user. Positive gap, though moderated because the shipping surfaces and Cowork trial figure are concrete.
Vendor staff on the record, competitive positioning against Anthropic
The primary sources are OpenAI employees whose product is the subject, speaking to promote a rollout and to assert that OpenAI has closed a gap with a competitor that shipped first. OpenAI simultaneously withheld training-data specifics, so the flow of information is selective. The publisher's incentive is access to a marquee lab's internal team. The one independent researcher partially offsets this.
Facts of the rollout are solid; capability and durability are not
High confidence that the surfaces shipped, that screen recording is part of setup, and that the reported Cowork trial figure was published as stated. Low confidence on relative capability, on whether trial converts to sustained use, and on the operational questions the dek raises, since one vendor-sourced article cannot settle them.
build
Codex learns to click: the coding agent stops typing patches and starts operating the machine1 distinct publisher
build
Developer habit, priced at $965B: what Anthropic's run actually proves1 distinct publisher
product
Record, don't prompt: two labs converge on demonstration as the agent interface1 distinct publisher
build
Claude Code now outruns Copilot roughly two to one in JetBrains' survey of 15,000 developers1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 20, 2026