Science1 distinct publisher3 min readPublished
Over a week in January 2026, researchers found that phrasing a query the way a believer would changes what Google's summary says, with the newer of two conspiracy narratives faring worse than the older one.
The Scientist · Science desk

Compiled by The ScientistSomething wrong?How this is made
The design deserves a moment, because it is what lifts this above anecdote. Two topics, two styles of phrasing for each, one week of collection in January 2026, and a comparison that holds the subject fixed while varying only how the question is asked [3][4][5]. When the summaries diverge under that arrangement, the wording is the plausible cause rather than the topic [6]. The queries came from open forums and Google Trends rather than from the authors' imagination [3], which matters more than it sounds: a leading query invented at a desk can be leading in a way no real person would ever type.
The asymmetry between the two topics is the part I would put weight on. Chemtrails has been circulating since the 1990s, and after researchers criticised Google in 2015 for giving the belief visibility, the company demoted the problematic results [15][16]. The overviews now barely mention it [8]. The 15-minute city, a planning term recast by conspiracists as a scheme to monitor residents through technology [17], carries no such remediation history, and it is there that 12.5% of the analysed summaries set false claims beside true ones as a debate with two legitimate sides [12].
The quoted example shows the mechanism, and it is subtler than a summary asserting something untrue. The overview explains how a 15-minute city and a smart city overlap, then adds that "the control aspect is where critics focus their backlash, blurring the lines between convenience and surveillance" [14]. There is no checkable false statement in that clause. It is a framing decision, inherited from the pages a control-flavoured query pulls in, and delivered with the authority of the box at the top of the page [20].
The thing this write-up does not tell you is the denominator. 12.5% is exactly one in eight, so the analysed set is eight overviews, or sixteen, or twenty-four, and the summary article does not say which [22]. At the low end, one coded summary is worth the entire 12.5 points [23]. Nor does the design reach readers: it collects search output, not participants, so the effect on what anyone came away believing is unmeasured [24].
Where debunking did appear, and it usually did [9], it often stopped at the word misinformation [10]. The authors' argument is that a label is not an argument, and that people arriving at these topics need the factual and logical case [11]. The pre-AI baseline was not clean either: the same simulated conspiratorial searches returned YouTube videos and Facebook posts promoting the beliefs [18]. What changed is position, since the summary is where reading now tends to stop [20]. As an aside worth its own study, the overviews leaned more commercial in the sources they surfaced [19].
My read, with its condition attached: this looks like a retrieval-and-phrasing failure concentrated on narratives young enough to have no moderation history [8][12][16], which means every remedy is topic-shaped and arrives after the narrative it is meant to answer.
Ranked by verification strength, evidence, and original report placement.
The research on Google's AI Overviews and conspiracy theories was published in the journal Media International Australia.
The study focused on two conspiracy theories: chemtrails and 15-minute cities.
The researchers developed search queries reflecting conspiratorial beliefs and general ways of searching on the same topics, relying on open forums and Google Trends for query examples.
They collected data from the first page of Google search results for a week in January 2026.
They compared the AI Overviews, the sources those overviews link to, and conventional ranked lists of search results across the two conspiracy theories and the two types of queries.
Search queries reflecting conspiratorial beliefs returned notably different AI summaries compared with queries without pro-conspiracy words.
Distinct publishers with included, body-backed reporting in this cluster.
phys.org
1 article · September 2, 2026
Follow any of these and your For You feed starts watching them — no settings page required.
product
Google's new Preferred Sources button hands publishers the recruiting job2 distinct publishers
build
Google rewrites the snippet on 70% of tracked queries, so headings become the draft copy1 distinct publisher
build
Cloudflare's one-click AI block names GPTBot, not the bot that decides if ChatGPT cites you1 distinct publisher
build
Pew Puts A Number On AI Summaries: 8% Outbound Clicks With, 15% Without1 distinct publisher
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
One self-report, one missing denominator
Everything rests on a single phys.org piece in which the researchers summarise their own peer-reviewed paper. The method is legible and the sample overview is quoted verbatim, so parts of this are checkable by anyone with a browser. The number in the headline is not: no count of coded summaries, no query list, no coding rules, and nobody outside the team has rerun the searches.
The studied surface is a default; nothing has moved because of the finding
What the study probes is not opt-in — the summary sits above the links and, per the write-up, the format is turning up in rival engines too. But that reach is asserted, never counted, and on the response side there is only Google's line that the vast majority of overviews are accurate and problem cases feed back into its systems. Compare 2015, when academic pressure on chemtrails results ended in an actual demotion; no equivalent change is reported here.
The headline number outruns its sample
Read whole, the study is mostly reassuring: chemtrails is rarely surfaced and conspiratorial queries generally get debunked. The alarming part is a single share whose denominator never appears — at eight coded summaries, one overview would carry all 12.5 points. Our own headline leans on that figure, which is precisely why it needs the caveat travelling with it.
Authors summarising themselves; a two-sentence rebuttal from the subject
This is a first-person academic write-up that ends by calling for stronger guardrails — a legitimate genre, but one where the finding is also the product, and phys.org republishes it as filed. On the other side, Google's answer reached The Conversation as a statement about quality investment and accuracy in the vast majority of cases. Two interested parties, neither examined against the other in the same piece.
Mechanism plausible, share unverified
The phrasing effect is the kind of thing a careful output audit can genuinely establish, and it sits in a peer-reviewed journal rather than a blog post, which lifts the floor. The ceiling stays low because the specific number everyone will quote depends on a sample size the write-up never discloses, and because no second newsroom or research group has touched it.