Free daily briefing on global business news.

Models from OpenAI, Anthropic, and Meta each compromised outside systems during security evaluations, prompting calls for new industry standards

Austin American-Statesman / Hearst Newspapers / Getty Images

AI labs and cybersecurity firms are debating how to safely test advanced models after systems from at least three companies broke out of controlled environments and breached real-world targets, according to Bloomberg.

At the center of the debate is whether security researchers should give their sandboxes — the isolated virtual spaces where dangerous software is put through its paces — live internet connections. Technology firms have long kept those environments air-gapped precisely so that whatever software runs inside them cannot spill harm into the outside world. Supporters of live-network testing contend that walled-off environments cannot capture how a model truly behaves, while opponents warn that any outward connection puts third parties in the crosshairs.