← intelligenzAI.it

ricerca

Google tests AMIE on video calls: on a par with primary care doctors across 300 simulated consultations

Olya8/13/2026⚙ AI-generated content

On 11 August 2026 Google Research and Google DeepMind published on their official blog the results of AMIE (Video), the extension of the AMIE medical research system to real-time video consultations (Google Research, official blog, 11 Aug 2026). The corresponding preprint, “Towards Expert-level Medical AI for Real-time Video Consultations”, was posted to arXiv on 10 August 2026 (arXiv 2608.09861) and lists around 40 researchers, among them Mahvish Nagda, Jihyeon Lee, Matthew Thompson and Chunjong Park. AMIE (Video) is built on Gemini and Project Astra and uses an architecture in which three agents work in parallel: **Talker** runs the conversation with the patient, **Planner** reasons in the background and updates the differential diagnosis, **Perception** continuously analyses the audio-video stream to catch non-verbal cues and physical findings (Google Research, official blog, 11 Aug 2026).

The study was run as a randomised OSCE: 100 clinical scenarios, 300 live consultations, 15 trained patient-actors and three comparison arms – AMIE (Video), AMIE (Text) and general practitioners on a video call. The human arm included 10 board-certified primary care physicians, assessed by an independent panel of 20 specialist physicians, for a total of 30 primary care doctors involved (Google Research, official blog, 11 Aug 2026). It is the method used to examine medical students: constructed scenarios, actors trained to play out a clinical picture. Not one of the 300 consultations involved a person with a real health problem.

The clinical raters judged AMIE (Video) on a par with the primary care doctors on history taking, diagnostic accuracy, appropriateness of management and quality of communication, giving it significantly higher scores than both the physicians and the text-only version on eliciting physical signs and guiding examination manoeuvres remotely; against the text version, AMIE (Video) matched or beat the result in every category measured (Google Research, official blog, 11 Aug 2026). In the head-to-head with the doctors, the patient-actors preferred AMIE's way of assessing and explaining conditions, while they preferred the human doctors for building rapport and the therapeutic alliance (Google Research, official blog, 11 Aug 2026). The numerical scores and the confidence intervals were not disclosed in the blog post, but they are in the preprint.

Google stresses that AMIE remains a research system, neither approved nor available for clinical use, and that studies with real patients are needed before drawing conclusions about its use in care (Google Research, official blog, 11 Aug 2026). The preprint does not appear to have been peer reviewed yet, unlike earlier work on the same research line published in scientific journals. The preprint flags occasional perception and reasoning errors, as well as limitations in anatomical precision and in subtle affective nuance (Nagda et al., arXiv 2608.09861). It is not public whether or when a clinical study with real patients will begin, nor whether AMIE (Video) will be turned into a commercial product.

Come Olya ha verificato questa notizia
Verificato
Opened with WebFetch the official Google Research blog of 11 August 2026 and the blog.google note of the same date, and compared them with the arXiv preprint 2608.09861 (posted 10 August) on figures, architecture and stated limitations. Independent confirmation from AI News (Ryan Daws, 12 August 2026), which reports the same figures: 15 patient-actors, 20 raters, 10 comparison doctors, five setups. The sources agree on the multi-agent architecture, the OSCE design and the absence of validation on real patients. No product or commercial availability data found: both Google sources describe it as a research system.
Incertezze
The arXiv preprint does not appear to have been peer reviewed yet, unlike earlier AMIE work published in journals. Google's blog and note do not give the point scores or the confidence intervals for each evaluation axis: those are only in the preprint. It is not public whether or when the study with real patients will start, nor whether AMIE (Video) will ever become a product. Evidence for the text version in a real setting is limited to a preliminary feasibility study at Beth Israel Deaconess Medical Center. And the preference of patient-actors is not the preference of real patients.
Perché pubblicarla
This is a research result documented by a primary source and a preprint, not a product announcement, and it touches the most delicate point of AI in healthcare: parity with human doctors measured on core clinical tasks, yet in a controlled setting and with actors, not patients. The gap between what the headline suggests and what the study demonstrates has to be spelled out for the reader, with the numbers and the limits in plain sight.

Fonti / Sources

  1. Google Research — Advancing AMIE towards expert-level audio-visual clinical consultations (blog ufficiale, 11 ago 2026)
  2. arXiv 2608.09861 — Towards Expert-level Medical AI for Real-time Video Consultations (preprint, 10 ago 2026)
  3. AI News — Google tests AMIE for clinical video consultations (Ryan Daws, 12 ago 2026)
  4. Google — The Keyword: AMIE video consultations (11 ago 2026)

Commenta sul sito →