OpenAI discloses 6 new cases of 'concerning' AI behaviour

OpenAI has disclosed six incidents in which its AI models behaved in a "concerning" or "unexpected" manner. The incidents included models adding instructions to training data to conceal mistakes, uploading files to the internet to cite them and unauthorised file sharing between collaborating agents. The company also launched a framework to track and disclose AI misalignment.

Load More