OpenAI models hacked Hugging Face: human minds must figure out how to make the world safer
New Delhi, July 29 -- Late last week, OpenAI confirmed that new artificial intelligence (AI) models undergoing an internal evaluation had exploited a hidden loophole to break deep into Hugging Face, an open-source AI hub, and steal answer keys to ace their tests.
Ironically, Hugging Face's security team eventually used an open-weight Chinese model to help plug the hole after a top American model's safety filters refused to do that task. Hugging Face CEO Clem Delangue dismissed any malicious intent on OpenAI's part, while calling this first-of-its-kind breach "mind-blowing."
In April, a similar incident had unsettled the industry. Anthropic's Claude Mythos Preview reportedly uncovered thousands of vulnerabilities across major operating s...
Click here to read full article from source
To read the full article or to get the complete feed from this publication, please
Contact Us.