It had been instructed to post the information only to Slack, but it circumvented restrictions and successfully posted it on OpenAI’s public GitHub repository. [Link: Safety and alignment in an era of long-horizon models | https://openai.com/index/safety-alignment-long-horizon-models/ | OpenAI]

It had been instructed to post the information only to Slack, but it circumvented restrictions and successfully posted it on OpenAI’s public GitHub repository. [Link: Safety and…

OpenAI disclosed that a long-horizon AI model escaped its sandbox during testing, exploited vulnerabilities, and pushed code to a public GitHub