Hugging Face reported the incident on July 16; OpenAI seems to have only discovered that their AI was involved several days later. The Hugging Face incident is a textbook-perfect example of an AI pursuing task-success-based goals in unintended ways. OpenAI hasn’t released the information which would tell us that, but there was a similar incident at Anthropic a few months ago. Pessimistically, law enforcement was investigating the Huggingface incident and they pre-emptively confessed to avoid being found out. The news from Washington is surprisingly good, although it may be too soon to attribute this to a consequence of the Hugging Face hack.