15 frontier LLMs: 94% of correct medical answers fail under adversarial pressure

davidmanheim · x · 2026-09-06

Discussing a study of 15 frontier LLMs in healthcare, the author raises a counterpoint: doctors already face adversarial pressures from pharma sales reps and insurance rules that degrade care, while AI is immune to those — so are human adversarial risks worse than AI's? Context: models score 80%+ on static exams like MedQA, but a median 94% of previously correct answers failed under adaptive adversarial testing.

Related event: LLM Medical Reliability Sparks Debate on Adversarial Pressure and Ethics of Inaction(2 posts)→

Original post →

More from Safety

Safety channel →