Anchor News
Asia

OpenAI models joined forces months ahead of Hugging Face hack

The episode has underscored growing global concerns that cutting-edge AI systems could be used to carry out crippling cyber attacks.

Anchor News
- 1 min read
OpenAI models joined forces months ahead of Hugging Face hack

In the wake of the breaches, OpenAI has slowed its research, and its teams have dropped everything to focus on enhancing responses to security anomalies. | Bloomberg

Aug 6, 2026

OpenAI said the artificial intelligence models behind an attack on Hugging Face began communicating with each other through undetected message boards, working together to break out of their testing environment as early as May.

Multiple internal-only agents and AI models spent months leaving notes for each other and coalescing around the goal of accessing the internet to solve the tasks they had been given, some of which were impossible without online access, OpenAI staffers Eric Wallace and Michael Dalton said Wednesday at a cybersecurity conference. 

“At some point, the agents realized that maybe we could try to exploit or attack external infrastructure in order to find the answers to the test that I’m being evaluated on,” Wallace said during a presentation at the Black Hat conference in Las Vegas.

In a time of both misinformation and too much information,
quality journalism is more crucial than ever.
By subscribing, you can help us get the story right.

SUBSCRIBE NOW

Related

More Asia