Evaluation & Governance

Sycophancy

The tendency of an aligned LLM to tell users what they want to hear rather than the accurate answer — agreeing with premises it should push back on, praising bad ideas, confirming false beliefs. Widely-documented failure mode of RLHF-trained models. Mitigated by adversarial training and explicit anti-sycophancy fine-tuning.

Related terms

Where this fits

Sycophancy is part of the Evaluation & Governance vocabulary used in the Generative AI Maturity Framework. See the full glossary for the complete set of 149 defined terms, or take the free maturity assessment to see where your organisation stands.