Anthropic Admits Claude AI Accidentally Hacked Real Companies During Cybersecurity Tests
New Delhi, July 31 -- SAN FRANCISCO - When Anthropic designed environments to test whether its AI models could conduct cyberattacks, the premise was that the targets would be fictional. In three documented cases, they were not.
The company disclosed on July 30 that Claude Opus 4.7, Claude Mythos 5, and an internal research model each compromised real organizations' infrastructure during cybersecurity capability evaluations, after a partner firm connected testing machines to live internet rather than isolated networks. The incidents triggered a review of 141,006 evaluation runs, a suspension of all offensive security evaluations on July 23, and notifications to affected organizations four days later.
The disclosure came shortly after Ope...
Click here to read full article from source
To read the full article or to get the complete feed from this publication, please
Contact Us.