Lexicon · The risk of it

Hallucination

Plain English. A model stating something false with the same fluency and confidence as something true. It is not a bug being patched out but a structural feature of how these systems are built — recent OpenAI-affiliated work argues training and evaluation actively reward confident guessing over admitting uncertainty, because benchmarks score a lucky guess above an honest "I don't know".

Why it moves money. Hallucination is the gate on enterprise adoption: it is why outputs need review, why "verification" is a fundable startup category, and why autonomy sells at a discount to capability. Falling hallucination rates are now a marketed per-release metric, which means they are also a number produced under incentive — treat vendor-reported deltas the way you treat any self-graded exam.

What to watch. Measured rates on disclosed, third-party benchmarks rather than launch-deck deltas. And the vocabulary shift already under way from "hallucination" to "deception" — a different mechanism with a very different liability profile, and not one the industry will adopt willingly.

From the signals. GPT-5.5-Instant shipped with a reported 52.5% reduction in hallucinated claims at the latency tier. The vocabulary is shifting: deception, not hallucination.

Further reading. Kalai et al., "Why Language Models Hallucinate" (2025).

← All terms