Because the cost of a wrong answer to an objection is a deal that generates no signal anywhere in your business. Nobody visits, nobody bounces, nobody opens a ticket: a person asked whether your category gets accounts banned, received a confident wrong answer, and went somewhere else. Every other prompt class fails softly by comparison. On a “what is X” prompt, an error is cosmetic and gets corrected the moment the buyer reads two more sources. On “is X safe to use”, the error ends the process.
The probability of a wrong answer is not small, and the honest way to state it is as a range with its samples attached. The Tow Center for Digital Journalism’s 2025 audit of 1,600 queries across eight engines found more than 60% returned incorrect answers, with the worst engine wrong 94% of the time and the best still wrong 37% of the time; its earlier test of 200 quotes found 153 responses partially or entirely incorrect, with uncertainty signalled 7 times.1 At the sentence level the picture is better but not comfortable: an academic study found only 51.5% of sentences in generative-search answers fully supported by their citations, and a 2026 study classified about 11% of 98,020 atomic claims as unsupported by the sources cited for them.23 Different denominators, different tasks, same conclusion: a citation is not a guarantee of an accurate description.
Those studies measured news attribution and claim support rather than product descriptions, so what they establish is not how often engines misdescribe your pricing but that the failure mode is common enough to plan for.