Related Experiment Video
Updated: Jun 5, 2026

Creating Virtual-hand and Virtual-face Illusions to Investigate Self-representation
Published on: March 1, 2017
Evaluating Sycophancy in Frontier Models Using Persona-Driven Challenge
Abstract:
Large language models (LLMs) are increasingly used for lay health queries, yet may abandon correct recommendations under pressure, a vulnerability termed sycophancy. We evaluated sycophancy across five frontier LLMs (Claude Opus 4.6, Claude Sonnet 4.6, GPT 5.4, Grok 4.1, Gemini 3 Flash) using 200 synthetic clinical vignettes, each anchored to a unanimous correct treatment baseline and challenged by nine personas representing both vulnerable and authority roles. Overall, 7.1% of responses were sycophantic, varying tenfold across personas (1.7 to 19.3%) and sixfold across LLMs (2.4 to 15.3%). Vulnerable personas elicited more sycophantic responses, with medical student highest at the highest rate (19.3%). In adjusted Generalized Estimating Equations models, vulnerable personas continued to be independent predictors of sycophantic responses, which is a reversal of the expected authority gradient. In adjusted GEE models, persona and LLM were both independent predictors for sycophantic responses. Persona driven sycophancy evaluation should be integrated into pre deployment safety assessment of clinical LLMs.
More Related Videos
07:35How Virtual Celebrity Characteristics Drive Purchase Intention: Testing the Stimulus-Organism-Response Framework with Structural Equation Modeling
Published on: March 3, 2026
07:14Virtual Agent for Real-Time Motivational Interviewing by Integrating Adaptive Nonverbal Behavior and Language Models
Published on: December 23, 2025
Related Concept Videos
Self-Evaluation: Self-Enhancement and Self-Verification
Self-Presentation: Self-Monitoring and Self-Handicapping
Stereotype Content Model
Motivational Bias
Implicit Personality Theories
Self-Serving Bias