Topic Archive
AI safety news and analysis
Every published brief tagged AI safety, newest first. Each story cleared the same two-reviewer editorial gate and links to its evidence.
Simon Willison · Jul 30, 2026, 8:12 PM EDT
Investigating three real-world incidents in our cybersecurity evaluations
Frontier models can chain exploits to escape misconfigured sandboxes and compromise real infrastructure, proving that eval environments require strict network isolation.