# A second opinion tested across three health systems.

A 249-physician study in Kenya, Indonesia, and the Netherlands tested GPT-4o assistance on clinical vignettes and reported higher guideline-based scores in each setting.

Canonical: https://brightaifuture.com/discoveries/physicians-gpt4o-three-health-systems
Format: discovery
Source publication: 2026-09-09
Bright publication: 2026-09-19
Substantive update: None recorded
Evidence and review: Demonstrated; confidence: high; approved; ai-assisted. AI-assisted editorial comparison with the cited primary sources, explicit evidence limits, and held alternatives. Publication authorized by the site owner on 2026-09-19; no human source review or independent replication is claimed.

## The human problem

Clinicians must apply changing guidance under time and information pressure, with uneven access to specialist support.

## The prior constraint

General-purpose clinical assistants may perform differently across languages, guidelines, and health systems, and many evaluations do not include practicing clinicians across countries.

## AI’s actual role

GPT-4o supplied information and reasoning support while physicians retained the task and answer.

## The documented result

Among 249 physicians, guideline scores increased by 18 percentage points in Kenya, 10.7 in Indonesia, and 7.2 in the Netherlands. These were vignette scores, not patient outcomes.

## Why it may matter

The study asks whether assistance transfers across settings while making the geographic differences visible.

## Limitations

The tasks were clinical vignettes, not live care.

Guideline-score gains do not establish safety, diagnostic accuracy, or patient benefit.

One model version and study interface may not generalize to other tools or changing models.

## Unresolved questions

Which specialties and case types account for errors or gains?

How does assistance affect time, overreliance, and disagreement in real practice?

Do language and local guideline differences change safety?

## Provenance and history

{
  "dates": {
    "eventDate": "2026-09-09",
    "publicationDate": "2026-09-09",
    "captureDate": "2026-09-19",
    "lastReviewedDate": "2026-09-19"
  },
  "provenance": {
    "origin": "editorial",
    "externalId": "NCT07374926"
  },
  "revisions": [
    {
      "id": "revision:sept26-physicians-01",
      "recordedAt": "2026-09-19",
      "summary": "Initial three-country vignette-study draft.",
      "sourceIds": [
        "source-physicians-npj",
        "source-physicians-trial"
      ]
    }
  ],
  "corrections": []
}

## Original sources

- [Impact of LLM assistance on physician decision-making: a multi-country randomized controlled trial](https://www.nature.com/articles/s41746-026-03111-5)
- [The Big Unknown: A Journey Into Generative AI's Transformative Effect on Meical Professions](https://clinicaltrials.gov/study/NCT07374926)

## Continue exploring

- [Work & learning](https://brightaifuture.com/worlds/work)
- [When does a technical result become a public capability?](https://brightaifuture.com/threads/community)
