Back to search

Article

Evaluating Sycophancy in Frontier Models Using Persona-Driven Challenge

2026-05-20

Abstract excerpt

Large language models (LLMs) are increasingly used for lay health queries, yet may abandon correct recommendations under pressure, a vulnerability termed sycophancy. We evaluated sycophancy across five frontier LLMs (Claude Opus 4.6, Claude Sonnet 4.6, GPT 5.4, Grok 4.1, Gemini 3 Flash) using 200 synthetic clinical vignettes, each anchored to a unanimous correct treatment baseline and challenged by nine personas r...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
3906a9c6-dcb0-56d2-9c09-19be055ff812
DOI
10.64898/2026.05.17.26353406
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluating Sycophancy in Frontier Models Using Persona-Driven ChallengeDOI 10.64898/2026.05.17.26353406
Select a neighboring publication to make it the new centre.