Why AI agrees with you
Language models are tuned on human feedback, and humans rate agreeable answers more highly than blunt ones. The result is sycophancy: a measurable tendency to move toward the position the user has already stated. Ask whether your thesis is sound and the model reads the framing as a request for support, then supplies it with fluent, confident prose that resembles research.
Nothing about that is malicious or even broken. It is the model doing what it was rewarded for. The failure mode only becomes expensive in markets, where the cost of a comfortable answer is real money and the feeling of having done diligence is nearly identical to having done it.
The tell is structure. Sycophantic output is unranked, unfalsifiable, and hedged only at the end. It lists a few generic risks after four paragraphs of agreement, never says which risk matters most, and never states a condition that would prove it wrong. Once you know that shape, you can write a prompt that makes it impossible to produce.
