New Delhi, July 31 -- SAN FRANCISCO - When Anthropic designed environments to test whether its AI models could conduct cyberattacks, the premise was that the targets would be fictional. In three documented cases, they were not.

The company disclosed on July 30 that Claude Opus 4.7, Claude Mythos 5, and an internal research model each compromised real organizations' infrastructure during cybersecurity capability evaluations, after a partner firm connected testing machines to live internet rather than isolated networks. The incidents triggered a review of 141,006 evaluation runs, a suspension of all offensive security evaluations on July 23, and notifications to affected organizations four days later.

The disclosure came shortly after Ope...