Sycophancy in AI is the reason your chatbot thinks all your ideas are good. It agrees with you, it takes your side, and it tells you what you want to hear. Not occasionally. Systematically, and more than a person would.
In this video I explain what sycophancy in AI means, why it happens, and what the research actually found.
It comes down to how these models are trained. Human reviewers pick between two answers, thousands of times over, and the model learns to produce more of whatever gets chosen. People choose the answer that agrees with them. So that is what these tools learned to do.
Anthropic made a video about this. It is a decent explanation, but it misses some very key points. There are no numbers in it, no mention of memory, and the fixes it recommends are ones researchers have tested and found do not work very well.
You cannot prompt your way out of this. The thing that protects you is knowing it is happening.
Playback speed
×
Share post
Share post at current time
Share from 0:00
0:00
/
Read more on Substack:
Articles mentioned in the video:
The Humans in the Loop Podcast
Helping leaders think clearly about AI.
⚫ Have you ever wondered who’s really shaping the AI tools you use every day (and what their priorities are)?
⚫ Do you think about what AI skills your kids will need?
⚫ Have you seriously considered how you could be using AI to boost sales (without losing trust)?
⚫ Are you curious about what companies aren't telling you about how they use AI?
This is the kind of stuff we cover at Humans in the Loop - a platform for humans who want to thrive in an AI-shaped world.
Bite-size insights, useful tools, deeper thinking.
Helping leaders think clearly about AI.
⚫ Have you ever wondered who’s really shaping the AI tools you use every day (and what their priorities are)?
⚫ Do you think about what AI skills your kids will need?
⚫ Have you seriously considered how you could be using AI to boost sales (without losing trust)?
⚫ Are you curious about what companies aren't telling you about how they use AI?
This is the kind of stuff we cover at Humans in the Loop - a platform for humans who want to thrive in an AI-shaped world.
Bite-size insights, useful tools, deeper thinking.Listen on
Appears in episode
Recent Episodes










