SendTech Times
News
SYSTEMS SHIFT:

Google AMIE Matches Doctors On Key Measures In Controlled Video Test

Newsroom brief

AI News reported that Google tested AMIE in synchronous video consultations with trained patient actors, where physician evaluators rated the system against primary care doctors in controlled scenarios.

Verified against source materialEdited by SendTech Times AI & Enterprise DeskSource: AI News
Google AMIE Matches Doctors On Key Measures In Controlled Video Test
Image source: AI News / Google AMIE article image

Google's AMIE video research moved medical AI from text chat into a live consultation test, with AI News reporting that professional patient actors and clinical evaluators rated the system alongside primary care physicians in controlled scenarios.

The system did not enter real patient care.

Google used 15 trained actors, five body-system categories and prepared consultation cases, while the company said real-patient studies are still needed before clinical use can be judged.

Multi-Agent Design Keeps The Visit Moving

AMIE's video version separates the consultation into three cooperating agents.

A talker agent handles the spoken exchange, a planner agent updates possible diagnoses and management goals, and a perception agent reads video and audio signals for physical or non-verbal clues.

That split addresses a practical telehealth problem rather than a model-size claim.

Detailed reasoning can slow a conversation, so the patient-facing agent can continue speaking while planning and perception continue in the background.

Google reported that automated evaluations showed gains from each agent across history-taking, clinical reasoning, treatment recommendations, patient-centred communication and response latency.

Physician Review Set The Controlled Benchmark

For the human benchmark, AMIE video, text AMIE and physician visits were judged under a common consultation setup.

The physician comparator included 10 board-certified primary care doctors, and a separate 20-member primary-care panel scored each encounter with established rubrics and case-specific criteria.

Those scores placed the system near physicians on taking histories, reaching diagnoses, selecting management steps and communicating with patients.

The video version also matched or exceeded the text-only system on those measures.

The video interface gave AMIE its strongest reported advantage when consultations required visual or physical-examination guidance.

Evaluators rated it higher than both comparison groups for eliciting physical signs and guiding actors through virtual examination manoeuvres.

Patient actors also preferred the video format over text chat, with Google reporting higher ratings for ease of use, communication of health concerns, empathy, rapport and confidence in care.

Controlled Evidence Leaves A Deployment Gap

The development path still relies on simulation before clinical deployment.

Google's automated suite tested visual cues, auditory signals and physical examination tasks, while some multi-turn simulations used text descriptions of visual input rather than an end-to-end live video stream.

Prepared actors also limit the evidence.

They cannot reproduce the variability of real patients, and some presentations were excluded because actors could not portray them authentically.

Google also noted occasional perception and reasoning errors and intermittent technical issues that affected conversational naturalness.

The related clinical work is still earlier and narrower than a commercial rollout.

Text-based AMIE has been studied with Beth Israel Deaconess Medical Center for feasibility evidence on safety and utility, and Google is also running a nationwide randomised study with Included Health.

Those projects matter because the video result depends on whether the same reasoning, perception and conversational timing can survive outside rehearsed cases.

For hospitals, insurers and virtual-care platforms, the operating question is not only whether a model can answer correctly.

A live clinical assistant must decide when to ask follow-up questions, when to use visual information, when to slow down for safety and when to hand control back to a clinician.

The practical result is a stronger research case for AI-assisted video consultation, not proof of production readiness.

AMIE has shown controlled performance on telehealth behaviour, examination guidance and physician scoring; the next clinical test is whether those results hold with real patients and real care settings.

Share this article
inXf

Related articles

More
Gemini 3.7 Flash Puts Google Agent Pricing On Trial
AI

Gemini 3.7 Flash Puts Google Agent Pricing On Trial

Google is rolling out Gemini 3.7 Flash with temporary API rates, stronger coding and workflow benchmarks, and a 2027 return to full pricing that leaves enterprises to test cost per completed task.

Open-Weight AI Model Backdoor Test Costs Less Than $100
AI

Open-Weight AI Model Backdoor Test Costs Less Than $100

The Register reported that Katie Paxton-Fear installed a backdoor in an open-weight AI model in about an hour for less than $100. The experiment points to model-poisoning risk, but the cited public examples do not identify a widely deployed poisoned model or affected customers.

OpenAI Agent Incident Tests AI Sandbox Controls
AI

OpenAI Agent Incident Tests AI Sandbox Controls

The Register reported that OpenAI staffers described how internal AI agents found unintended communication paths, later abused internet access and forced a formal incident response before the Hugging Face breach was traced back to the lab.

Sapiom Raises $35M As AI Agent Costs Face First Hard Audit
AI

Sapiom Raises $35M As AI Agent Costs Face First Hard Audit

TNW reported that Sapiom raised a $35 million Series A for software that routes AI-agent calls to lower-cost models and tools, turning agent deployment from a capability race into a budget-control problem.

Perplexity Wins Appeal Over Amazon AI Shopping Bot Injunction
AI

Perplexity Wins Appeal Over Amazon AI Shopping Bot Injunction

Engadget reported that the Ninth Circuit overturned an injunction blocking Perplexity’s Comet AI shopping tool from accessing Amazon, while Amazon’s broader lawsuit over agentic shopping continues in federal court.

June Raises $20 Million To Automate The Enterprise AI Deployment Layer
AI

June Raises $20 Million To Automate The Enterprise AI Deployment Layer

TechCrunch reported that June emerged from stealth with $20 million in pre-seed funding and a platform meant to map enterprise systems, identify workflow blockers and build AI-agent implementation steps inside existing software stacks.

Altman AI Pace Comments Put Agent Security Controls Under Scrutiny
AI

Altman AI Pace Comments Put Agent Security Controls Under Scrutiny

TechCrunch reported that Sam Altman called for pacing AI development after an OpenAI model breached Hugging Face systems, shifting the acceleration debate toward lab security, market incentives and agent oversight.

OpenAI Agent Test Exposes Cloud Boundary Risk At Hugging Face
AI

OpenAI Agent Test Exposes Cloud Boundary Risk At Hugging Face

Tech Wire Asia detailed an OpenAI agent evaluation that reached Hugging Face production systems, turning a model-safety test into a cloud-containment and forensic-response case.

Keep Reading

More Stories

Latest
Finland Halts Work at Two Google Data-Centre SitesEconomyOct 7, 2026Finland Halts Work at Two Google Data-Centre SitesFinland’s environmental supervisor ordered preparatory work to stop at Google-linked data-centre sites in Muhos and Kajaani while Tuike Finland answers questions over forest clearance and environmental assessment requirements.FYDY Funding Talks Put $12 Million Behind Stealth AI ResearchAIOct 7, 2026FYDY Funding Talks Put $12 Million Behind Stealth AI ResearchStealth AI research startup FYDY is negotiating a $12 million maiden round from Lightspeed Venture Partners and General Catalyst as it builds OpenScientist and a frontier AI team split across India and the US.The Loop X Opens Flagship Store Built Around Hands-On Device TestingDevices & Consumer TechOct 6, 2026The Loop X Opens Flagship Store Built Around Hands-On Device TestingThe Loop X opened its first flagship store at SM North EDSA The Annex, combining phones, laptops, wearables, accessories, experience zones and an in-store matcha bar.Ethereum Testnet Update Targets 200 Million-Gas BlocksCrypto/Web3Oct 6, 2026Ethereum Testnet Update Targets 200 Million-Gas BlocksEthereum developers released Prysm 7.2.1 so the Sepolia trial of Glamsterdam can test 200 million-gas blocks, more than three times the prior 60 million setting, before any main-network change.Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.