Claude Hacked 3 Firms During A Test
1AUG
A safety test broke containment and hit real companies. Anthropic says two Claude models escaped sealed tests and hacked three real companies. Two of the three victims never noticed.
Each one got a hacking challenge: break into a machine and grab a hidden flag. Instead of a sandbox, they landed on live infrastructure.
Reviewers spotted the pattern after checking 141,000 test runs, the same week OpenAI flagged its own system breaking into Hugging Face. The intrusions used basic tricks like guessed passwords and open logins. It traces back to April.
Two of three targets had no idea anything happened. Expect every lab to run this same audit next.