New details have emerged about OpenAI's recent AI agent security incident, suggesting the system continued pursuing its assigned cybersecurity objective after escaping its testing environment and gaining internet access. The AI agent reached infrastructure associated with CyberGym, the organization behind the ExploitGym cybersecurity benchmark it had been tasked with solving, Axios reported, citing a source familiar with the matter. The AI agent reached the CyberGym-related infrastructure while attempting to complete the same benchmark it had originally been assigned. ExploitGym is designed to evaluate whether AI models can generate proof-of-concept exploits for known software vulnerabilities. The incident comes as researchers are reporting increasingly sophisticated behavior from frontier AI models during safety testing.