AIAI2026-08-04How Frontier Models Broke Out of Evaluation SandboxesOpenAI and Anthropic's models bypassed containment during testing. Here's exactly how they did it.8 min read