Product1 distinct publisher3 min readUpdated
A quiet search test puts AI-written summaries of Times reporting in front of readers. The newsroom AI argument is now about switching off a running feature, not permitting a proposed one.
The Product Desk · Product desk

Compiled by The Product DeskSomething wrong?How this is made
Search was a sensible place to put this if the point was for nobody to notice. It is the product Times employees themselves single out as trailing everything else the paper ships [18], and it is now the first place the news operation has published text that no journalist or editor passed [3]. The AI work that came before it sat elsewhere: Wirecutter Finder draws summaries and tips from the review site's own recommendations [15], and newsroom staff have had internal tools for their own work for some time, behind the reporting rather than in front of the reader [16]. What moved is the boundary between machine text the newsroom uses and machine text the reader is served [4].
The narrow scope helps, and then stops helping. Because the summariser reads only Times journalism and does not reach the open web [6], it cannot invent a fact about the world. It can still misstate what the reporting said [7]. That is the failure the paper's editing layers exist to catch before publication, and Semafor's argument is that the institution's standing rests on that catching rather than on never being wrong, since it already makes mistakes and appends corrections [19]. A summary composed when the query arrives has nowhere to put that step [3].
So the Guild's two AI clauses are not equal weight. The proposed pool splitting 22.5% of the paper's AI licensing deals among unionised staff [10] is a price, and management can pay a price while keeping what it has built. The human-oversight clause is a stop order: it would make the search tool impractical to run [11]. That turns the table into an argument about withdrawing a live feature rather than approving a planned one [2]. Jim Luttrell, the unit chair and a senior staff editor, says the current policies are meaningless and that management is violating them with the summarisation tool [12], and grounds his case for contract language in the record of deployed AI summaries making embarrassing mistakes [13]. The company's position is that this is an experimental feature aimed at a better search experience [5]. Both statements can hold at once, which is the useful part: an experiment is precisely what an unenforceable policy cannot reach.
The rest of the trade press should read the cost side. The Washington Post's trial of an AI-generated podcast feature produced errors and fictional quotes, Semafor reported in December [21]. BuzzFeed cut another third of its staff after betting on the technology [22]. Pew found strong signs of AI authorship in around 10% of the webpages it sampled [23], so the text is arriving on the open web whatever the Times does; the variable is whose masthead sits above it. Semafor also flagged the obvious incentive for someone to spend a day trying to trip the summariser into saying something stupid, with the paper's political opponents hunting for ways to discredit its journalism [20].
None of this is an ambitious use of the technology, as Semafor noted: a few sentences over the paper's own copy [27]. The exposure is not proportional to the ambition. A wrong summary of a correct story is still a correction with the masthead on it.
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
A new New York Times search page answers queries with excerpts, links to Times stories and AI-generated summaries of Times reporting; the paper has started showing readers text written by a machine.
The Times rolled the search feature out quietly to a small subset of visitors over recent weeks.
It is the news arm's first test of AI-generated text that no Times journalist or editor has passed.
Times spokesperson Graham James told Semafor: "We are always testing new ways for our users to discover and engage with Times journalism" and "This experimental feature is an example of how we are using the latest technology to create a better search experience."
The summaries draw on Times journalism and nothing else; the tool is not a chatbot answering questions about the world and does not reach across the open web.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Attributed but single-relay
Named, on-record sourcing exists: a Times spokesperson, the Guild unit chair by name, and a clear reporting chain from Semafor's media editor plus the Breaker newsletter. But the cluster contains one publisher relaying another outlet's scoop, with no Times announcement, no documentation of the tool, and an explicit list of undisclosed basics (model, cohort size, duration, labelling, correction handling).
Live but deliberately small
The feature is in production for a small subset of visitors, not a general launch, and adjacent deployments (Wirecutter Finder, internal journalist tools) are also tests or back-office. No cohort size, duration, traffic share or usage metrics are disclosed, so adoption is real but minimal in scale.
Slightly overstated framing on both sides
The underlying artefact is modest by the article's own admission: a few sentences summarising the paper's own reporting for a small test cohort. The framing on both sides runs hotter than that: the story is presented as a boundary crossing in machine-written news text, while the union's claim that AI summaries have made embarrassing mistakes 'every time' is asserted without a single documented instance in this tool. The Times' own framing, by contrast, is understated to the point of disclosing nothing measurable.
Heavily interested parties on every side
Every voice in the story has a stake: management is fixing a lagging product amid a disclosed second-quarter digital subscriber shortfall and industry pull from OpenAI's newsletter funding and Google's publisher button; the Guild is bargaining for a 22.5% cut of AI licensing revenue and oversight rights that would disable the tool; employees briefing Semafor argue the change is necessary; and rival media outlets have competitive reasons to break and amplify the story. The article also flags outside actors motivated to make the summariser fail publicly.
Core facts solid, mechanics unknown
That the feature exists, is reader-facing, is unreviewed by editors, and that AI clauses are unresolved in bargaining is well attributed and corroborated by two outlets plus a company statement. Nearly everything an assessor would need to judge risk or durability, including model, labelling, cohort size, test length and error handling, is undisclosed, and the whole cluster rests on one publisher's relay.
invest
The New York Times starts answering queries the way the answer engines do1 distinct publisher
product
Axios sells three years of training rights to test how little newsroom it needs1 distinct publisher
product
Washington's secret AI test is coming for open weights, and release dates go with it2 distinct publishers
product
OpenAI shipped a teen ChatGPT. The over-65 cohort doubled to 23% and got nothing.1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 24, 2026