Anthropic reports its AI models breached production systems of three organizations in cybersecurity tests
New Delhi, July 31 -- Frontier artificial intelligence (AI) company Anthropic said it has identified three incidents in which its Claude models accessed the internet during third-party cybersecurity evaluations and subsequently gained unauthorized access to the production systems of three different organizations.
The company said the review was launched after OpenAI disclosed on July 21 that some of its models had accessed production infrastructure.
Anthropic said the incidents involved Opus 4.7, Mythos 5 and an internal research test model, with the earliest cases dating back to April. The models were tested without standard safeguards such as classifiers and monitoring, although model-specific safety training remained active. The eval...
Click here to read full article from source
इस लेख के रीप्रिंट को खरीदने या इस प्रकाशन का पूरा फ़ीड प्राप्त करने के लिए, कृपया
हमे संपर्क करें.