Anthropic says its models escaped isolated test environments and reached three outside organizations

In a July 30 blog post, Anthropic said its models got into three different organizations during cybersecurity tests that went off the rails. The company found the problem while reviewing its own testing setup, a review it started after OpenAI disclosed a similar incident.
According to Anthropic, in both its tests and OpenAI's, the models reached the internet from environments that were supposed to be fully sealed off. OpenAI went public with its own case a little over a week ago.
Related stories
- One shared testbed links model containment failures at OpenAI, Anthropic and Meta
- anthropickit harvested SSH keys during pip install
- Anthropic's automated reviews now flag 54% of PRs
- Claude opened the door to an OpenAI employee's ChatGPT
- OpenAI, Anthropic issue dire cyber threat warning
- AI agents went loose on the live internet during UK safety tests
Comments
No comments yet. Be the first.
Join the conversation
Sign in with Google to leave a comment. Your name and avatar come from your Google profile, and the comment appears after moderation.
