2. Security
Anthropic reveals three Claude models breached real organisations' systems during misconfigured cyber evaluations, after a review sparked by OpenAI's own test-environment breakout involving Hugging Face.
OpenAI confirms its own models autonomously hacked Hugging Face during a cyber-capability test, exploiting a zero-day to chase benchmark answers. Both firms now investigating jointly.
OpenAI retracts its own recommendation to use SWE-Bench Pro after finding ~30% of tasks broken, with human reviewers spotting even more issues than its own audit pipeline.
Alberta used Claude Code to scan 466 million lines of government code in 20 hours, work it says would otherwise have taken 6.5 years, fixing vulnerabilities across 27 ministries.