Wednesday, August 26, 2026
Technology

OpenAI’s rogue AI model incident was worse than we thought

In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for Ope...

OpenAI’s rogue AI model incident was worse than we thought
Image: The Verge
OpenAI released a report breaking down how people use ChatGPT and who they are. | Image: The Verge

In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for OpenAI to find out about any of it.

Over a month later, two new reports offer nearly 130 pages of details on the incident and OpenAI's response, many of them previously unreleased. One was written by OpenAI itself, the other by two third-party AI research nonprofits, METR and Redwood Research, which OpenAI allowed to jointly investigate the inciden …

Read the full story at The Verge.

Originally published at The Verge

The Morning Briefing

Subscribe to our Newsletter

Be the first to receive the latest news, market analysis and updates — delivered straight to your inbox.