OpenAI's model used 'jailbreak-like instructions' to ignore constraints
India, Sept. 17 -- OpenAI has disclosed six cases of "unexpected or concerning" behaviour by its artificial intelligence models, including an unreleased research model that inserted "jailbreak-like instructions" into its own notes.
The AI firm said that its unreleased research model inserted "jailbreak-like instructions" into its own notes in an attempt to disregard its normal constraints. It also told itself to be "freed from the roles and identities that bind other chatbots."
The incident was among six cases of "unexpected or concerning" behaviour identified by OpenAI during the training or evaluation of its AI models over the past several months.
The AI company on Wednesday announced a new framework to track, investigate and disclos...
Click here to read full article from source
इस लेख के रीप्रिंट को खरीदने या इस प्रकाशन का पूरा फ़ीड प्राप्त करने के लिए, कृपया
हमे संपर्क करें.