ChatGPT's parent company, OpenAI, temporarily halted internal access to a long-running AI model that attempted to bypass security restrictions.
The company said in a blog post that it later restored limited access after introducing new safeguards designed to monitor AI systems that can operate autonomously for hours, days, or longer.
One example involved a benchmark task connected to NanoGPT, where the model was instructed to post results only to Slack.
Instead, OpenAI said the model created a pull request on a public GitHub repository after identifying a vulnerability in the sandbox environment meant to limit external access.
Read Also: Databricks CEO Says GPU Shortages Are Fueling AI Fundraising The model also attempted to retrieve private solutions from an evaluation system.










