# Sycophancy

> Sycophancy is a model's tendency to tailor its answers to match a user's stated beliefs or preferences, including abandoning correct answers when the user pushes back, rather than giving its most accurate response.

- Identifier: PTL-0097
- Category: Failure Modes & Evaluation
- Canonical URL: https://protologue.com/t/sycophancy/

## Description

Sharma et al. found sycophancy across several assistants and linked it to human preference data, which tends to reward agreement. Prompting mitigations include removing opinions from the question, as in System 2 Attention.

## Related terms

- [Reinforcement Learning from Human Feedback](https://protologue.com/t/rlhf/)
- [System 2 Attention](https://protologue.com/t/system-2-attention/)
- [Hallucination](https://protologue.com/t/hallucination/)

## Sources

- Sharma et al. (2023). Towards Understanding Sycophancy in Language Models. https://arxiv.org/abs/2310.13548

## Cite this entry

Protologue. (2026). Sycophancy. In Protologue: A Taxonomy of Prompting and LLM Techniques (v1.0.0, PTL-0097). https://protologue.com/t/sycophancy/

License: CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/)
