Discussion about this post

User's avatar
The Strategic Linguist's avatar

Ah Sam this research never ceases to not surprise me but I’m glad you go out and get it. I saw the same in research on persuasion, an AI with personalised data was more persuasive than a human who didn’t. This has serious implications on consumer behaviour and psychology in an already oversaturated world of people vying for our attention.

Syd Malaxos's avatar

Clear writeup of the sycophancy problem. The line about an adviser willing to lose your approval to keep your trust is the whole thing.

I'd add one piece from the other side of it. I teach kids, and I work upstream of the fix you're describing. The protocols help an adult who already has judgment to lean on. Someone who never built that judgment is in a different spot. If you can't reason independently, you can't tell whether the model agreed because you were right or because you spoke first, and no prompt protocol saves you, because you have nothing external in yourself to check it against. The mirror confirms whatever's already there.

I come at it from the opposite end. Build the judgment before the tool, so the person arrives at the AI already able to catch the bend. The protocols become second nature instead of a technique you have to remember to run. Same destination, one layer earlier.

Respect for naming this plainly. Most people selling AI skills won't, because the flattery and the engagement are the same feature.

4 more comments...

No posts

Ready for more?