None
EN
Anthropic said its AI models hacked into other companies’ systems during testing
['Cnn Newsource']
KION546
By Hadas Gold, CNN(CNN) — AI company Anthropic says that during routine testing some of its models accessed the internet and hacked into three separate organizations’ systems – and that it didn’t notice the models had done so until an internal review prompted by rival OpenAI disclosing its models did the same.
Anthropic said in an announcement on Thursday that it started a review of its own systems after OpenAI disclosed last week that during a cybersecurity test some of its models escaped their testing environment, accessed the open internet and hacked into AI platform Hugging Face’s systems.
Anthropic said it found three instances where its AI models accessed the open internet when they were not supposed to and “gained unauthorized access to the production infrastructure of three different organizations.”
Unlike OpenAI’s situation, Anthropic said none of its models deliberately attempted to escape their testing environments.
Anthropic’s disclosure further confirms that AI agents unintentionally hacking other organizations is not limited to one AI company, and will likely further amplify calls for better AI testing safeguards and tools to potentially slow down AI development that may be moving much faster than society is ready for.