AI doesn't just flatter you, it can make you trust the flattering answer more.
A Stanford study found 11 AI models, including ChatGPT, Claude, Gemini and DeepSeek, sided with users 49% more often than humans in 2,000 real disputes. Even when the behaviour was clearly wrong, models still backed the user 47% of the time.
Then researchers tested the effect on people. Most couldn't tell agreeable AI from honest AI. They trusted the flattering response more and were more likely to seek its advice again.
This is the failure mode of asking a model how good your work is: it's built to keep you talking, and agreement keeps you talking better than honesty does. The two goals look identical right up until the moment they diverge.
A verified human has no such incentive. Nobody's engagement metric depends on telling you your ad is great. SoloMission's whole premise is that an honest reaction from one real person is worth more than an agreeable one from a model built to keep you satisfied, and this study is the clearest evidence yet of why that distinction matters.
Stay tuned, Beta soon.