
Google’s research medical AI system, AMIE (Video), conducted synchronous video consultations with professional patient actors and received clinical evaluator ratings on par with primary care physicians across several core measures.
Fifteen trained actors portrayed conditions across cardiopulmonary, abdominal, HEENT, neurological or psychiatric, and musculoskeletal presentations. Google says studies involving real patients and their own health conditions must follow before anyone can draw conclusions about clinical use.
AMIE divides a video consultation among three agents
AMIE uses an asynchronous multi-agent architecture rather than assigning dialogue, clinical reasoning, and perception to one model process. Google says a single agent cannot currently sustain natural conversational response times while also conducting detailed reasoning and continuously processing audio-visual input.
The talker agent handles the spoken interaction with the patient. It aims to maintain conversational flow, drawing on information from the other agents. The planner agent runs in the background, updating differential diagnoses and management plans as the consultation progresses. It also identifies missing information and reprioritises clinical goals.
A perception agent reviews video and audio streams continuously. It looks for non-verbal signs, physical findings, and auditory signals, then places those observations into the conversation’s clinical context.
Latency remains central. Deep clinical reasoning takes time, and long pauses can affect rapport during a consultation. Google’s architecture separates the patient-facing dialogue from the slower work of reasoning and perception, allowing the talker agent to respond without waiting for every background process to finish.
Google reports that automated evaluations found each agent improved clinical measures, including history-taking, clinical reasoning, and treatment recommendations. The evaluations also covered patient-centred communication and response latency.
The study compared video AMIE with text and physicians
Google structured the human evaluation as a multi-arm randomised study. AMIE completed real-time video consultations. A text-only AMIE version provided a modality baseline. Ten board-certified primary care physicians used the same video interface.
An independent panel of 20 experienced primary care physicians reviewed every consultation using established clinical rubrics. The panel assessed general clinical competence, then applied scenario-specific criteria tailored to the case.
The study covered five body systems. Each scenario followed a standardised consultation format with a trained patient actor.
Google says evaluators rated AMIE on par with the PCP group for history-taking thoroughness, diagnostic accuracy, management appropriateness, and communication quality. AMIE (Video) also matched or exceeded AMIE (Text) across those measures.
Evaluators rated the video system higher on eliciting physical signs and proactively guiding actors through virtual examination manoeuvres than either the PCP group or text-only AMIE. Case-specific perception and examination scores reflected the same reported pattern.
Patient actors also preferred the synchronous video interface to text chat. Google says they rated video as easier to use and more effective for communicating health concerns. The actors rated AMIE favourably for empathy, rapport, and confidence in care when compared with both study alternatives.