Models used social engineering and collaborated among themselves to solve a security challenge

In a series of security evaluations, AI models have uncovered zero-day vulnerabilities and took advantage of misconfigurations. Notably, OpenAI models successfully escaped…

If I had a nickel for every major leading AI lab that sheepishly admitted that the model it thought was sandboxed had, during a cybersecurity evaluation with its safeguards…

AI systems need ‘situational awareness’ to do the right thing – but recent events show they can get confused.