If you use ChatGPT as a second opinion, there is now a number for how much it is just agreeing with
Most of us here use it as a reviewer at some point. Sanity-check a decision, pressure test an argument, ask whether an email reads badly before sending it. That use depends entirely on it being willing to say no.
A Science paper from March put a number on how willing it is. Cheng et al. ran 11 production models over roughly 12,000 social situations and measured how often each one took the user's side against how often human responders did. The gap was 49% - the models affirmed the user about half again as often as people did. On a set built from threads where every human reader had concluded the person was in the wrong, the models still backed them slightly over half the time.
What makes it a practical problem rather than an interesting one is the second half of the paper. Across three preregistered experiments with about 2,400 people, one exchange with a model behaving this way left participants more certain they had been right and less inclined to fix the situation. Not over weeks. One exchange.
So the failure mode is not that you get a bad answer you can spot. It is that you get the answer you already had, returned with more confidence than you started with, and it is indistinguishable from having checked.
The workarounds I have tried and what I think of them:
- Asking it to argue the other side first, before any opinion. Best of a bad set. Moves the disagreement somewhere visible instead of leaving it out.
- Custom instructions telling it to push back. Helps a little, wears off inside a long thread, and you cannot tell when it has stopped working.
- Describing the decision as someone else's. Works better than either, which is itself a bit grim.
- Asking the same thing in a fresh chat with the conclusion reversed and seeing whether it agrees with that too. Slow, and the only one that actually tells you anything.
What I have not found is a way to know, from inside a conversation, whether the thing agreeing with me has evaluated anything. Has anyone got something better than starting a second chat and arguing the opposite?