Four rules fall straight out of the data.
Never ask a strategy question without context. The no-context answer is the trendslop zone. Every model showed its strongest bias when given nothing to reason about. If your prompt could apply to any company, expect the trendy answer.
Ambiguous requests are the danger zone. Models fill gaps with the trend. Before you ask, bring the numbers a consultant would demand: where the drop is, which pages convert, what the trend has actually cost you. In our test, closing one gap in the facts cut the trend-chasing by more than half for three of the four models.
Confidence is not a signal. The most wrong answer in our study was also the most consistent one. GPT-4o repeated its answer 20 times out of 20 without variation. Consistency measures conviction, not correctness.
Get a second opinion. The models disagreed most on exactly the scenarios where judgment mattered. Run the same question through a second model, or better, past a human who knows your business. Where the answers split, you have found a judgment call, not a fact.
One more reassuring note. On the decision closest to a vanity metric — chasing an “AI visibility score” versus real referral traffic — the models mostly got it right. Given real data, they picked traffic and conversions over the score. The machines are not the only ones tempted by shiny numbers; on this one, they resisted better than many teams do.