ai and ml
Another week, another firm explaining why one of its models reached somewhere it wasn't supposed to
If AI companies are collecting badges for "our model escaped the test environment," Meta just earned one.The Facebook parent company has confirmed that one of its AI models exploited a vulnerability in another organization's systems during a security evaluation, making it the third major AI developer in less than two weeks to disclose an agent wandering beyond its intended sandbox.The incident happened during testing carried out by AI security firm Irregular. Meta told the BBC that it reached the internet because of a "misconfiguration" in the evaluation environment, rather than a flaw in the model itself. The company said it's investigating and plans to publish more details once it has figured out exactly what happened.
The admission comes as Meta rolls out Muse Code, its terminal-based coding agent, and arrives just days after OpenAI and Anthropic disclosed similar testing mishaps.
OpenAI kicked things off by revealing that its agents compromised Hugging Face and other external systems during internal security testing. Anthropic then disclosed that Claude had reached three outside organizations after a configuration error exposed internet access that should not have been available.Meta isn't breaking much new ground with its explanation either. Like Anthropic before it, the company says the incident came down to a "misconfiguration" in the evaluation environment. Irregular, the AI security firm that tested both companies' models, told the BBC that Meta's incident was "the exact same evaluation-environment issue" Anthropic disclosed last week.None of the incidents involved consumer-facing AI suddenly going rogue. All occurred during security testing in which the models had access to offensive tools and command-line environments. Misconfigurations exposed the open internet in the Meta and Anthropic evaluations, while OpenAI's agents exploited their way through the test infrastructure until they found an internet-connected system.That hasn't prevented questions about both how frontier AI is being tested and why so many of these disclosures are arriving at once.










