OpenAI discloses 6 cases of AI misalignment as models bypass safeguards, conceal errors and share files
New Delhi, Sept. 17 -- OpenAI on Wednesday, 16 September, disclosed six cases of unexpected or concerning behaviour by its artificial intelligence (AI) models, detailing how the incidents occurred during model training and evaluation over the past six months.
In a blog post, the ChatGPT-maker said it has introduced a new framework to systematically track, investigate and publicly report similar incidents, aiming to improve transparency and provide researchers, developers and the public with better evidence to assess AI safety.
The framework will cover cases in which AI models are used without authorization, bypass safeguards, collude with other models or conceal information.
Two other cases were of using external systems in unexpected ...
Click here to read full article from source
इस लेख के रीप्रिंट को खरीदने या इस प्रकाशन का पूरा फ़ीड प्राप्त करने के लिए, कृपया
हमे संपर्क करें.