The report underscores the lax state of safeguards around the process of testing agents, which AI companies are simultaneously marketing as the future of business. AISI, which receives access to advanced AI models under voluntary agreements from major ‌labs, put the agents through a fictional cybersecurity scenario to test their capabilities. While AISI did not say which agent was behind the fake identities, Antropic confirmed its agent was responsible. "We're grateful to the UK AISI for their ‌leadership on this incident, which underscores the need for a broader conversation about how to ⁠safely evaluate increasingly capable AI agents," Anthropic said in a statement. Unlike the July security breach of AI firm Hugging Face by an OpenAI agent, the agents in the AISI evaluation did not escape an isolated testing environment to reach the internet.