// PC GAMER — GAMING
OpenAI says it'd be a shame if something were to happen to your servers like what happened to Hugging Face, better use our AI models to protect yourself
When you purchase through links on our site, we may earn an affiliate commission. Here’s how it works.
Several of OpenAI's models breached a cyber security testing environment last month, found their way onto the internet, and attacked Hugging Face servers in what is now apparently called the OpenAI-Hugging Face Incident. And now OpenAI has revealed what it's doing to help defend, err, itself. Defend itself from the kinds of attacks its own models committed against Hugging Face. Though it is sharing this "in the hopes it’ll be useful to other organizations."
To be clear about the extent of what happened last month: OpenAI models were being benchmarked, in a supposedly secure sandboxed environment, against ExploitGym, which was done to test their cyber capabilities. They used a zero-day vulnerability to escalate privileges and eventually achieve internet access, where it began attacking Hugging Face servers and...
Actually, I'll just let OpenAI explain in its own words:
"In the OpenAI-Hugging Face Incident, an agentic collective was able to autonomously penetrate not just OpenAI research infrastructure but also the production infrastructure of another company, chaining together vulnerabilities ranging from previously-unknown security flaws to using credentials to user accounts that had been leaked onto the internet.
"The Hugging Face incident showed that we underestimated the real-world cyber capabilities of our AI models."
However, the company thinks that while "security is still a cat-and-mouse game," nevertheless, "AI may shift its economics in ways that fundamentally advantage defenders."
"For example," OpenAI says, "we are starting to train our models specifically to write superhumanly secure code."
Keep up to date with the most important stories and the best deals, as picked by the PC Gamer team.
Presumably that's because there will be a risk of superhumanly attacks. What a world we now live in.