Build1 publisher3 min readPublished
A reader test written up by dev.to shows Gemini's search tool being withheld before the model runs, and because the reply carries no marker either way, a stale answer is indistinguishable from a searched one.
The Engineer · Build desk
Compiled by The EngineerSomething wrong?How this is made
A grounded answer and a memory answer arrive in the same shape. Nothing in the Gemini reply separates them, and there is no field a caller can read to tell which one turned up [1]. Withholding a tool per turn is cheap and defensible engineering, and the dev.to write-up grants that case in full [13]; the interface choice is to publish nothing about the decision.
The test that pins the mechanism is well built, and worth reading for the construction. Weather works as a probe because, per that write-up, Gemini has no separate weather feed, so current conditions come through the same search tool as everything else [5]. Hold the account and the settings fixed, ask in English, and it searches. Ask the same thing with the single Finnish word "saa" and search is blocked, on every attempt [4]. In the visible reasoning, the model works out that it should check current conditions, then notes it has not been handed the tool [6]. One variable moved, and it was the language, which puts the deciding step before inference [1]. dev.to reads the result the same way, as a routing decision made before the model starts rather than a limit of the model [2].
The leaked system prompt is the weaker leg. The line is verbatim, "Do NOT issue search queries to the google search tool for this prompt", in a repository the article describes as public, heavily starred and covered by mainstream press, and the author states plainly that they cannot audit Google's internal prompts and do not claim to [7][8]. For that line to explain your session, it would have to be the prompt served to your account, on your surface, in your language, on that turn. What survives without the repository is the behaviour and the model's own account of it, including reports of Gemini reproducing the line back to users [9].
The conflict case is the one I would want logged. On some turns the model is handed the search tool, an instruction to always use it, and the instruction not to issue search queries, all at once, and users have watched it referee that in its chain of thought, at one point wondering whether one of the instructions is a prompt injection [10]. A classifier that resolves its own contradictions inside the model's reasoning has moved the decision to the one place you cannot query.
Everything above is the consumer assistant: help threads named for Android and Web [3], a Reddit report from late August, and a complaint the article says has persisted across app updates through August 2026 [12]. The material says nothing about the developer API, so none of it is evidence about API grounding parameters. Google has published no explanation, no switch and no timeline [11]. Until there is a marker in the response, the paired-language probe is the only instrument on offer, and it needs re-running after each app release, since the behaviour has already outlived several [12].
Ranked by verification strength, evidence, and original report placement.
dev.to states that Google runs a routing step deciding question by question whether the model is handed its web-search tool, and that on many turns it either withholds the tool or attaches an instruction telling the model not to use it; the article calls the language-dependent result the signature of a routing decision made before the model starts, not a limitation of the model.
dev.to reports that Gemini frequently answers questions that obviously need the live web instantly and without searching, assembling the answer from training data that may be months or years stale, with no spinner, no sources and nothing in the reply indicating it never looked.
Gemini's official help community carries threads titled "Persistent Web Search Failure on Gemini Android and Web", "Google restricts Gemini to use web-search tool" and "Gemini refuses to use search tool", in which users describe questions that plainly need current information met with a stale answer and no search.
A user on r/GeminiAI in late August reported that asking Gemini for the weather in English produces a search, while asking the identical thing in Finnish with the single word "saa" blocks search every time, on the same account and the same settings.
Weather is a clean probe because Gemini has no separate weather feed: current conditions come through the same search tool as everything else, so the result is a straight check of whether search was handed over.
In the reported test the model's own visible reasoning works out that it should check current conditions, then notes that it has not been given the tool.
Follow any of these and your For You feed starts watching them — no settings page required.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
One write-up resting on user-generated corroboration
The sturdiest leg is public and checkable: three named threads on Google's own help community, quoted by title. Everything else is thinner. The central demonstration is one r/GeminiAI user's English-versus-Finnish weather comparison, and the documentary anchor is a leaked-prompt repository that dev.to quotes verbatim while saying plainly it has not audited Google's prompts. That candour improves the write-up's honesty without adding a second observer.
Deployment scope left unknown
This story has nothing countable in it: no version where the routing behaviour begins, no share of turns in which the search tool is withheld, no count of affected accounts or languages beyond English and Finnish. The one durable time reference — a complaint that survives app updates through August 2026 — speaks to how long the pattern has lasted, not how widely it lands.
Headline reaches further than one account can carry
dev.to's framing is "often", and the demonstration behind it is one language pair on one account. The gap stays small because the piece polices itself harder than most: it builds the strongest version of Google's routing argument, treats "the model got dumber" complaints as usually overstated, flags the unshipped toggle as one user's finding, and refuses to allege cost-cutting. What remains overstated is frequency, not mechanism.
A downside beat reporting a downside, unanswered by Google
The byline publishes under "The AI Downside" on a self-publishing developer platform, so the attention accrues to finding fault, and no editor outside the author gates the claim. Pulling the other way, the piece declines the motive that would have made the story bigger — that Google is throttling search to cut costs — and marks its unverified elements as unverified. The party holding the routing logs never speaks, which leaves the incentive imbalance uncorrected rather than exploited.
Mechanism coherent, breadth unestablished
Two independent-ish artefacts point the same way: the model's own reasoning saying it was not given the tool, and a leaked instruction telling it not to search. That coherence is why the mechanism reads plausibly. What it does not establish is how general the behaviour is, whether the leaked prompt is authentic, or whether Google would describe its routing this way, and no second publisher or vendor statement is available to close any of those.
build
A 5x publishing increase cost one site 1,000 indexed pages and every impression1 publisher
build
Before you spend quota on an agent skill, make it pass an eval harness1 publisher
build
Geofencing beats GPS polling on power, then loses to the OEM battery optimiser1 publisher
build
Google Trends returns 200 OK with an empty body when it blocks you1 publisher
Publishers with included, body-backed reporting in this cluster.
1 article · September 7, 2026