No, OpenAI’s models didn’t go ‘rogue’ when they broke into Hugging Face. Here’s what really happened.


When OpenAI recently revealed that two of its most advanced artificial intelligence (AI) models escaped cybersecurity testing and got hacked into a startup, looking a lot like the kind of scenario that AI security researchers have warned about for years.

The models found previously unknown vulnerabilities in the infrastructure built to contain them, gained access to the public internet, and broke into Hugging Face, a major platform for hosting AI models and datasets. However, their objective was less sinister than the sequence of events: they were looking for information that would help them pass a cybersecurity test given to them by OpenAI.

Leave a Comment