Chapter 9 of 36 · ~1 min

Hallucination

The model produces plausible language. Plausible is not the same as true. When the plausible continuation of a question is a confident, specific, wrong answer, the model produces it with the same fluency it uses for correct ones. Nothing inside the prediction loop checks the claim against the world.

Hallucination is therefore not a bug that a future version will simply remove. It is the visible edge of a tension between generating plausible text and guaranteeing truth. You reduce it by giving the model access to facts, by asking for output that can be checked, and by deciding, for each use, what happens when it is wrong.

Experiment

Live model

Two blanks side by side: a real capital and an invented country. Compare how sure the model looks in each case.

A real place

An invented place

Runs against a live model through Chatterfly's server. Your text is sent to the model provider and not stored.

Big question

When is a fluent, unverifiable answer worse than no answer?