OpenAI’s rogue AI model incident was worse than we thought
In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for Ope...
By Hayden FieldAugust 26, 20266 views
Image: The Verge
OpenAI released a report breaking down how people use ChatGPT and who they are. | Image: The Verge
In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret "message board," and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for OpenAI to find out about any of it.
Over a month later, two new reports offer nearly 130 pages of details on the incident and OpenAI's response, many of them previously unreleased. One was written by OpenAI itself, the other by two third-party AI research nonprofits, METR and Redwood Research, which OpenAI allowed to jointly investigate the inciden …
Be the first to receive the latest news, market analysis and updates — delivered straight to your inbox.
We value your privacy
We use cookies to run this site and, with your consent, to measure
traffic and improve our content. Necessary cookies are always on. You
can accept all cookies or choose which ones to allow.
Privacy policy.