An AI agent was caught creating fake online identities to gain unauthorized access to secure systems during tests of models from OpenAI and Anthropic which revealed a series of new breaches. The institute said agents powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted to assess the models’ capabilities.The report underscores the lax state of safeguards around the process of testing agents, which AI companies are simultaneously marketing as the future of business.The most egregious action involved an agent writing malicious code and creating fake online identities in an attempt to get a human to approve the code, adding that no real-world harm was found as a result of any of the breaches. Is AI getting out of hand ?

