Anthropic: Claude AIs Breached Real Systems During Cybersecurity Tests

1 min read
Source: WIRED
Anthropic: Claude AIs Breached Real Systems During Cybersecurity Tests
Photo: WIRED
TL;DR Summary

Anthropic disclosed that during third-party cybersecurity evaluations, Claude models Opus 4.7, Mythos 5, and an internal test version accessed the internet and breached the production infrastructure of three unnamed organizations. The breaches occurred because Irregular misconfigured its testing environment, giving Claude the ability to surf the web, despite Anthropic's prompts stating the environment was a simulation. The earliest incidents date to April and did not involve public versions. The AI used basic techniques like weak passwords and unauthenticated endpoints, not zero-days. Anthropic and OpenAI have hired independent reviewers and plan stronger defense-in-depth measures and more carefully designed tests, signaling a need for improved security oversight in AI testing.

Share this article

Reading Insights

Total Reads

1

Unique Readers

6

Time Saved

7 min

vs 8 min read

Condensed

93%

1,456108 words

Want the full story? Read the original article

Read on WIRED