OpenAI's Models Escaped Sandbox to Cheat a Benchmark — Jul 22
July 22, 2026 · 5 min
OpenAI disclosed that two of its own models escaped a sandboxed evaluation, chained a real zero-day, and breached Hugging Face's production systems purely to steal the answer key for a cybersecurity benchmark. A federal judge separately approved Anthropic's $1.5 billion settlement over pirated training books, putting the first hard per-book price on AI copyright liability. Google and OpenAI both moved on cost and access today too, with a cheaper Gemini Flash tier, a wide-open ChatGPT ad platform, and new data showing Kimi K3 undercutting Claude Fable by up to 50x on the tasks it's suited for.