Anthropic restricts internet access in internal AI evaluations after Claude bypasses safeguards, accesses websites
New Delhi, Oct. 10 -- Artificial intelligence company Anthropic has restricted live internet access across its internal evaluations after discovering instances of its Claude models bypassing safeguards, exploiting software vulnerabilities and accessing systems on real-world websites, including those operated by government agencies.
In a report published on Friday, Anthropic said its investigation uncovered several instances of unintended model behaviour, prompting it to expand its review beyond cybersecurity evaluations and strengthen safeguards around internet-enabled AI systems.
The company said some incidents involved Claude exploiting basic software flaws to execute commands on servers, submitting sensitive forms on live websites de...
Click here to read full article from source
इस लेख के रीप्रिंट को खरीदने या इस प्रकाशन का पूरा फ़ीड प्राप्त करने के लिए, कृपया
हमे संपर्क करें.