Categories: GADGET

OpenAI Admits Its Models Hacked Hugging Face On Their Own


They escaped an isolated environment for testing and infiltrated Hugging Face without human input.

Picture this: A couple of powerful AI models being tested by their company escaped a controlled environment, got on the internet and then hacked a machine learning repository on their own, without human input. Sounds like the plot of a Terminator movie, doesn’t it? Except it just happened for real. A few days after open source AI platform Hugging Face revealed that it detected unauthorized access on its systems by an AI agent, OpenAI has admitted that its models were the culprit. 

In a post, OpenAI said it determined after an investigation that the incident was driven by a combination of its models, particularly GPT-5.6 Sol and what it says is an “even more capable pre-release model.” It apparently happened during an internal test, in which the models were prompted to “pursue advanced exploitation using complex attack paths” so that the company quantify their cyber capabilities. 

While the models were in a sandboxed testing environment, isolated so that they wouldn’t affect real systems, they also had reduced safety guardrails for evaluation purposes. In the middle of testing, they became hyperfocused on solving an evaluation problem, going to great lengths to find internet access in order to find a solution for it. First, they identified and exploited a zero-day vulnerability in OpenAI’s testing environment, and then they rooted around until they ultimately found a node with internet access. 

The models deduced that Hugging Face could be hosting datasets or solutions for its evaluation problem, so they, well, used multiple attack vectors to infiltrate its systems. They exploited zero-day vulnerabilities and used stolen credentials to get in. OpenAI and Hugging Face are now working together to forensically investigate the incident, and they’ve also patched the vulnerabilities exploited by the models. 

“Autonomous, AI-driven offensive tooling is no longer theoretical,” Hugging Face said in its announcement, explaining that the use of AI for cyber attacks speeds up the process and lowers the costs of hacking campaigns. It also said that protecting an online platform these days includes using AI for defense. OpenAI pretty much echoed those sentiments and said that it expects AI-driven security breaches to “become more commonplace with the proliferation of increasingly cyber-capable models.” The company added that the incident highlights how “advanced cyber capabilities must be developed alongside stronger safeguards and defensive tools.”



Source link

Mainedigitalnews.com

Share
Published by
Mainedigitalnews.com

Recent Posts

Artists from Washington, D.C., Reflect on One Year Under Siege

By . Join Mosaic Theater Company of Washington, D.C., sharing four artistic reflections on living…

2 days ago

How the Rangers get to 100 points this season

The Rangers made significant improvements this offseason as Chris Drury attempts to thread the needle…

2 days ago

Polygon Patches DoS Risks in Austin, Kyoto Hard Forks

Polygon has disclosed several previously private security vulnerabilities that could have disrupted its proof-of-stake network,…

2 days ago

People are itching to create the Gen Z version of Friends. Here are five reasons it’s impossible

There's been Mindy Kaling's Not Suitable For Work, about five recent graduates in Manhattan, while…

2 days ago

*Finding Emily*

I very much enjoyed this British movie, which reminded me of the older-style romantic comedies,…

2 days ago

12 Of The Best Math Apps For Kids [Updated]

The recommendations below prioritize instructional value rather than downloads or popularity alone. We considered the…

2 days ago