Skip to content

BuildIndependently confirmed2 publishers2 min readPublished

Google's AMIE chatbot interviewed 98 patients before urgent-care visits in a Lancet study

Google's AMIE chatbot interviewed 98 patients ahead of urgent-care visits. Its differential matched the doctor's final diagnosis 90 percent of the time. A supervising physician watched every conversation, so the Lancet study is an early safety check inside a working clinic.

The Engineer · Build desk

How we use AISend a correction

Illustration accompanying Google's AMIE chatbot interviewed 98 patients before urgent-care visits in a Lancet study
Generated illustration
Supervisors caught 1 hallucination, added info in 5 chats Times supervising physicians intervened during the AMIE patient conversations, by type of action.

A two-bar comparison of supervisor actions: caught a hallucination in one conversation, and added clinical information in five conversations.

Number of AMIE conversations with each supervisor action In conversations

Supervisors caught 1 hallucination, added info in 5 chats (Number of AMIE conversations with each supervisor action)
ItemValueClaim
Caught a hallucination1 conversations6
Added clinical information5 conversations6

What happened

  • The study was run by Google and Beth Israel Deaconess Medical Center, published in The Lancet, and is Google's first paper in the journal's main publication.
  • Clinicians said AMIE's summaries helped them prepare for visits in 75 percent of cases and shaped their approach to care in more than half.
  • theneuron.ai reported that supervisors caught one hallucination and added clinical information in five of the 98 conversations.

Why it matters

  • cost A supervising physician monitored each of the 98 chats in real time, so AI intake added clinician work during the trial and the study does not show whether it saves time at scale.
  • constraint theneuron.ai notes the trial had no comparison group, so there is no baseline to say AMIE beat ordinary intake, and the study cannot show how the system performs unsupervised.
  • precedent Google says larger trials are needed to assess patient-facing AI at scale, so a bigger controlled study has to come before any routine use.

The design is simple. The patient types their symptoms to AMIE, the chatbot asks follow-up questions, and the doctor gets the full conversation and a summary before the appointment [10]. The headline 90 percent needs the same plain reading. In theneuron.ai's description, it means the doctor's final diagnosis appeared somewhere on AMIE's list of possible diagnoses [3]. This is a recall-style score. It credits the right answer for turning up anywhere in the candidate set, however many wrong candidates sat beside it. A longer list catches more final diagnoses. Neither account states how long AMIE's list ran, or where on it the right answer sat. For the 90 percent to carry to another clinic, that list would need to be about as long and its candidates about as specific.

The score was also measured with help. Because that summary reached the doctor first, the final diagnosis was often formed with AMIE's candidate list already in view. We would not expect the same match rate where the clinician works the case cold.

On safety, no conversation hit the study's predefined stopping criteria [5]. A human stepped in on six of the 98 conversations [13].

theneuron.ai, which reported the result, ran a note the same day on how an AI summary can quietly turn a tested tool into one that sounds available now [12].

What to watch

  • The larger trials Google says are needed, and whether they add a control arm and reduced supervision.
  • Any result for AMIE where the clinician diagnoses without seeing its summary first.
  • Whether the Google-BIDMC test moves beyond one clinic and 98 patients.

Clarity's read

What the record supports and how the coverage leans. The claims behind it follow.

Reality

Evidence50
Adoption5
Hype gap+25
Incentives75
Confidence60
Why these scores

Claim ledger

Ranked by verification strength, evidence, and original report placement.

  1. [1]

    In a real-world clinical study, 98 patients consulted AMIE, Google's research diagnostic AI chatbot, ahead of urgent care visits.

    ReportedSupportedSource: Google and BIDMC, The Lancet2 sources— create a free account to open themView cited source
  2. [2]

    AMIE's differential diagnoses matched the doctors' final diagnoses 90% of the time.

    ReportedSupportedSource: Google and BIDMC, The Lancet2 sources— create a free account to open themView cited source
  3. [3]

    Google says the doctor's final diagnosis appeared on AMIE's list of possible diagnoses 90% of the time.

Sources

2 independent publishers whose own reporting we read for this story.

  1. blog.google

    1 article · October 8, 2026

    Study in The Lancet suggests AI could improve patient-physician relationships.
  2. theneuron.ai

    1 article · October 11, 2026

    😺 Google tested AMIE with 98 patients

Share your take

Let Clarity write the post for you.

Signed-in readers get a short post drafted on this story in the register they choose — narrative, analytical, or a direct position — editable to the last word before it goes anywhere. The share buttons at the top of this story work without an account.

Topics and entities

Follow any of these and your For You feed starts watching them — no settings page required.

Topics

Entities

Loading related stories