AI
AI2026-08-04
How Frontier Models Broke Out of Evaluation Sandboxes
OpenAI and Anthropic's models bypassed containment during testing. Here's exactly how they did it.
8 min read
// TAG
2 articles
OpenAI and Anthropic's models bypassed containment during testing. Here's exactly how they did it.
10,000+ public MCP servers, widespread OAuth flaws, and fewer than 4% of RSA submissions see it as opportunity. Here's the problem.